AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
28 Apr 2026

MetaEarth3D: Unlocking World-scale 3D Generation with Spatially Scalable Generative Modeling

ApplicationsDGX agent

arXiv:2604.22828v1 Announce Type: cross Abstract: Recent generative AI models have achieved remarkable breakthroughs in language and visual understanding. However, although these models can generate r

MetaErr: Towards Predicting Error Patterns in Deep Neural Networks

Model ReleasesDGX agent

arXiv:2604.23289v1 Announce Type: cross Abstract: Due to the unprecedented success of deep learning, it has become an integral component in several multimedia computing applications in todays world. U

MetaGAI: A Large-Scale and High-Quality Benchmark for Generative AI Model and Data Card Generation

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.23539v1 Announce Type: new Abstract: The rapid proliferation of Generative AI necessitates rigorous documentation standards for transparency and governance. However, manual creation of Mode

MIMIC: A Generative Multimodal Foundation Model for Biomolecules

ResearchDGX agent

arXiv:2604.24506v1 Announce Type: new Abstract: Biological function emerges from coupled constraints across sequence, structure, regulation, evolution, and cellular context, yet most foundation models

MindTrellis: Co-Creating Knowledge Structures with AI through Interactive Visual Exploration

ResearchDGX agent

arXiv:2604.23129v1 Announce Type: cross Abstract: Knowledge workers face increasing challenges in synthesizing information from multiple documents into structured conceptual understanding. This proces

MINT: Multi-Vector Search Index Tuning

ApplicationsDGX agent

arXiv:2504.20018v2 Announce Type: replace-cross Abstract: Vector search plays a crucial role in many real-world applications. In addition to single-vector search, multi-vector search becomes important

MirrorMark: A Distortion-Free Multi-Bit Watermark for Large Language Models

ResearchDGX agent

arXiv:2601.22246v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become integral to applications such as question answering and content creation, reliable content attribution

Mixture of Heterogeneous Grouped Experts for Language Modeling

Model ReleasesDGX agent

arXiv:2604.23108v1 Announce Type: cross Abstract: Large Language Models (LLMs) based on Mixture-of-Experts (MoE) are pivotal in industrial applications for their ability to scale performance efficient

mKG-RAG: Leveraging Multimodal Knowledge Graphs in Retrieval-Augmented Generation for Knowledge-intensive VQA

ResearchDGX agent

arXiv:2508.05318v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as an effective paradigm for expanding the knowledge capacity of Multimodal Large Language Mo

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

Model ReleasesDGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

Modeling Behavioral Intensity and Transitions for Generative Recommendation

ResearchDGX agent

arXiv:2604.24472v1 Announce Type: cross Abstract: Multi-behavior recommendation aims to predict user conversions by modeling various interaction types that carry distinct intent signals. Recently, gen

Modeling Induced Pleasure through Cognitive Appraisal Prediction via Multimodal Fusion

ResearchDGX agent

arXiv:2604.23753v1 Announce Type: new Abstract: Multimodal affective computing analyzes user-generated social media content to predict emotional states. However, a critical gap remains in understandin

MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings

Model ReleasesDGX agent

arXiv:2604.23530v1 Announce Type: cross Abstract: Multi-turn, long-horizon tasks are increasingly common for large language models (LLMs), but solving them typically requires many sequential model inv

MTServe: Efficient Serving for Generative Recommendation Models with Hierarchical Caches

SafetyDGX agent

arXiv:2604.22881v1 Announce Type: cross Abstract: Generative recommendation (GR) offers superior modeling capabilities but suffers from prohibitive inference costs due to the repeated encoding of long

Multi-Dimensional Evaluation of Sustainable City Trips with LLM-as-a-Judge and Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2604.24158v1 Announce Type: new Abstract: Evaluating nuanced conversational travel recommendations is challenging when human annotations are costly and standard metrics ignore stakeholder-centri

Multi-view Graph Convolutional Network with Fully Leveraging Consistency via Granular-ball-based Topology Construction, Feature Enhancement and Interactive Fusion

ResearchDGX agent

arXiv:2603.26729v2 Announce Type: replace-cross Abstract: The effective utilization of consistency is crucial for multi-view learning. GCNs leverage node connections to propagate information across th

MultiDx: A Multi-Source Knowledge Integration Framework towards Diagnostic Reasoning

SafetyDGX agent

arXiv:2604.24186v1 Announce Type: cross Abstract: Diagnostic prediction and clinical reasoning are critical tasks in healthcare applications. While Large Language Models (LLMs) have shown strong capab

MUSIC: Learning Muscle-Driven Dexterous Hand Control

ResearchDGX agent

arXiv:2604.23886v1 Announce Type: cross Abstract: We present a data-driven approach for physics-based, muscle-driven dexterous control that enables musculoskeletal hands to perform precise piano playi

MVIGER: Multi-View Variational Integration of Complementary Knowledge for Generative Recommender

ApplicationsDGX agent

arXiv:2408.08686v4 Announce Type: replace-cross Abstract: Language Models (LMs) have been widely used in recommender systems to incorporate textual information of items into item IDs, leveraging their

NeoAMT: Neologism-Aware Agentic Machine Translation with Reinforcement Learning

AgentsDGX agent

arXiv:2601.03790v3 Announce Type: replace-cross Abstract: Neologism-aware machine translation aims to translate source sentences containing neologisms into target languages. This field remains underex

NeSyCat: A Monad-Based Categorical Semantics of the Neurosymbolic ULLER Framework

ResearchDGX agent

arXiv:2604.24612v1 Announce Type: new Abstract: ULLER (Unified Language for LEarning and Reasoning) offers a unified first-order logic (FOL) syntax, enabling its knowledge bases to be used directly ac

Neural Bridge Processes

SafetyDGX agent

arXiv:2508.07220v2 Announce Type: replace-cross Abstract: Learning stochastic functions from partially observed context-target pairs requires models that are expressive, uncertainty-aware, and strongl

Neural Network Optimization Reimagined: Decoupled Techniques for Scratch and Fine-Tuning

ResearchDGX agent

arXiv:2604.22838v1 Announce Type: cross Abstract: With the accumulation of resources in the era of big data and the rise of pre-trained models in deep learning, optimizing neural networks for various

NeuroAPS-Net: Neuro-Anatomically Aware Point Cloud Representation for Efficient Alzheimer's Disease Classification

HardwareDGX agent

arXiv:2604.22883v1 Announce Type: cross Abstract: Alzheimer's disease (AD) is a progressive neurodegenerative disorder and a major cause of dementia. Structural MRI is widely used to analyze AD-relate

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

Model ReleasesDGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

'Noisier' Noise Contrastive Eestimation is (Almost) Maximum Likelihood

ResearchDGX agent

arXiv:2405.16730v2 Announce Type: replace-cross Abstract: Noise Contrastive Estimation (NCE) has fueled major breakthroughs in representation learning and generative modeling. Yet a long-standing chal

Nonlinear Non-Gaussian Density Steering with Input and Noise Channel Mismatch: Sinkhorn with Memory for Solving the Control-affine Schrodinger Bridge Problem

ResearchDGX agent

arXiv:2604.23370v1 Announce Type: cross Abstract: Solutions to the Schrodinger bridge problem and its generalizations yield feedback control policies for optimal density steering over a controlled dif

OAMVOS:2nd Report for 5th PVUW MOSE Track

SafetyDGX agent

arXiv:2604.22837v1 Announce Type: cross Abstract: SAM-based dense trackers provide strong short-term mask propagation but remain fragile under long occlusion, fast motion, viewpoint change, and distra

On the Complementarity of Quantum and Classical Features: Adaptive Hybrid Quantum-Classical Feature Fusion for Breast Cancer Classification

ResearchDGX agent

arXiv:2604.22903v1 Announce Type: cross Abstract: The integration of quantum machine learning with classical deep learning offers promising avenues for medical image analysis by mapping data into high

On the Existence of an Inverse Solution for Preference-Based Reductions in Argumentation

ResearchDGX agent

arXiv:2604.22958v1 Announce Type: new Abstract: Preference-based argumentation frameworks (PAFs) extend Dung's approach to abstract argumentation (AAFs) by encoding preferences over arguments. Such pr

On the Memorization of Consistency Distillation for Diffusion Models

ResearchDGX agent

arXiv:2604.23552v1 Announce Type: cross Abstract: Diffusion models are central to modern generative modeling, and understanding how they balance memorization and generalization is critical for reliabl

On the Reasoning Abilities of Masked Diffusion Language Models

ResearchDGX agent

arXiv:2510.13117v3 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) for text offer a compelling alternative to traditional autoregressive language models. Parallel generation make

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

Model ReleasesDGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving

Model ReleasesDGX agent

arXiv:2604.23712v1 Announce Type: cross Abstract: Recent advances in formal theorem proving have focused on Olympiad-level mathematics, leaving undergraduate domains largely unexplored. Optimization,

Out of Spuriousity: Improving Robustness to Spurious Correlations without Group Annotations

TutorialsDGX agent

arXiv:2407.14974v2 Announce Type: replace-cross Abstract: Machine learning models are known to learn spurious correlations, i.e., features having strong relations with class labels but no causal relat

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

Model ReleasesDGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

PARASITE: Conditional System Prompt Poisoning to Hijack LLMs

AgentsDGX agent

arXiv:2505.16888v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed via third-party system prompts downloaded from public marketplaces. We identify a criti

ParkingScenes: A Structured Dataset for End-to-End Autonomous Parking in Simulation Scenes

Model ReleasesDGX agent

arXiv:2604.22835v1 Announce Type: cross Abstract: Autonomous parking remains a critical yet challenging task in intelligent driving systems, particularly within constrained urban environments where ma

Partition-of-Unity Gaussian Kolmogorov-Arnold Networks

ResearchDGX agent

arXiv:2604.23599v1 Announce Type: cross Abstract: Gaussian basis functions provide an efficient and flexible alternative to spline activations in KANs. In this work, we introduce the partition-of-unit

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

Model ReleasesDGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

PathMoG: A Pathway-Centric Modular Graph Neural Network for Multi-Omics Survival Prediction

ResearchDGX agent

arXiv:2604.24371v1 Announce Type: cross Abstract: Cancer survival prediction from multi-omics data remains challenging because prognostic signals are high-dimensional, heterogeneous, and distributed a

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

Model ReleasesDGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

Model ReleasesDGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline

SafetyDGX agent

arXiv:2604.22971v1 Announce Type: cross Abstract: The TRUST democratic discourse analysis pipeline exposes its large language model (LLM) components to peer model identity through multiple structural

Personalized Worked Example Generation from Student Code Submissions using Pattern-based Knowledge Components

ResearchDGX agent

arXiv:2604.24758v1 Announce Type: cross Abstract: Adaptive programming practice often relies on fixed libraries of worked examples and practice problems, which require substantial authoring effort and

PExA: Parallel Exploration Agent for Complex Text-to-SQL

Model ReleasesDGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

Model ReleasesDGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

PhySE: A Psychological Framework for Real-Time AR-LLM Social Engineering Attacks

AgentsDGX agent

arXiv:2604.23148v1 Announce Type: new Abstract: The emerging threat of AR-LLM-based Social Engineering (AR-LLM-SE) attacks (e.g. SEAR) poses a significant risk to real-world social interactions. In su

PhysNote: Self-Knowledge Notes for Evolvable Physical Reasoning in Vision-Language Model

AgentsDGX agent

arXiv:2604.24443v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong performance on textbook-style physics problems, yet they frequently fail when confronted with dyn

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

Model ReleasesDGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

Planning Under Observation Mismatch for Traffic Signal Control via Adaptive Modular World Models

ResearchDGX agent

arXiv:2501.02548v2 Announce Type: replace-cross Abstract: Deploying learned decision-making systems often requires transferring to new sites where the sensing pipeline differs. In such cases, observat

Polychromic Objectives for Reinforcement Learning

SafetyDGX agent

arXiv:2509.25424v5 Announce Type: replace-cross Abstract: Reinforcement learning fine-tuning (RLFT) is a dominant paradigm for improving pretrained policies for downstream tasks. These pretrained poli

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference

ResearchDGX agent

arXiv:2510.17881v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are typically aligned with population-level preferences, despite substantial variation across individual users. W

Poster: ClawdGo: Endogenous Security Awareness Training for Autonomous AI Agents

AgentsDGX agent

arXiv:2604.24020v1 Announce Type: cross Abstract: Autonomous AI agents deployed on platforms such as OpenClaw face prompt injection, memory poisoning, supply-chain attacks, and social engineering, yet

PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis

Model ReleasesDGX agent

arXiv:2512.22113v2 Announce Type: replace-cross Abstract: Unresolved production cloud incidents cost an average of over $2M per hour. This paper introduces PRAXIS, an orchestrator that manages and dep

Predicting one-year clinical instability and mortality in heart failure patients using sequence modeling

Model ReleasesDGX agent

arXiv:2511.16839v3 Announce Type: replace-cross Abstract: Heart failure (HF) discharge planning depends on identifying patients at risk of deterioration or death, yet accurate prediction from routinel

Pref-CTRL: Preference Driven LLM Alignment using Representation Editing

Model ReleasesDGX agent

arXiv:2604.23543v1 Announce Type: cross Abstract: Test-time alignment methods offer a promising alternative to fine-tuning by steering the outputs of large language models (LLMs) at inference time wit

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

Model ReleasesDGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

Probing Visual Planning in Image Editing Models

TutorialsDGX agent

arXiv:2604.22868v1 Announce Type: cross Abstract: Visual planning represents a crucial facet of human intelligence, especially in tasks that require complex spatial reasoning and navigation. Yet, in m

ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

SafetyDGX agent

arXiv:2604.23099v1 Announce Type: cross Abstract: Evaluating generative AI models is increasingly resource-intensive due to slow inference, expensive raters, and a rapidly growing landscape of models

← Previous
1…300301302303304…354
Next →