AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
14 May 2026

AgenticFlict: A Large-Scale Dataset of Merge Conflicts in AI Coding Agent Pull Requests on GitHub

AgentsDGX agent

arXiv:2604.03551v2 Announce Type: replace-cross Abstract: Software Engineering 3.0 marks a paradigm shift in software development, in which AI coding agents are no longer just assistive tools but acti

AgentLens: Revealing The Lucky Pass Problem in SWE-Agent Evaluation

AgentsDGX agent

arXiv:2605.12925v1 Announce Type: cross Abstract: Evaluation of software engineering (SWE) agents is dominated by a binary signal: whether the final patch passes the tests. This outcome-only view trea

AI co-mathematician: Accelerating mathematicians with agentic AI

AgentsDGX agent

arXiv:2605.06651v2 Announce Type: replace Abstract: We introduce the AI co-mathematician, a workbench for mathematicians to interactively leverage AI agents to pursue open-ended research. The AI co-ma


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AI-Generated Slides: Are They Good? Can Students Tell?

Model ReleasesDGX agent

arXiv:2605.13532v1 Announce Type: new Abstract: As generative AI (GenAI) tools become easily accessible, there is promise in using such tools to support instructors. To that end, this paper examines u

AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents

AgentsDGX agent

arXiv:2605.13357v1 Announce Type: cross Abstract: Foundation models have transformed automated code generation, yet autonomous software-engineering agents remain unreliable in realistic development se

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions

SafetyDGX agent

arXiv:2408.12935v4 Announce Type: replace Abstract: AI Safety is an emerging area of critical importance to the safe adoption and deployment of AI systems. With the rapid proliferation of AI and espec

AIvaluateXR: An Evaluation Framework for on-Device AI in XR with Benchmarking Results

Local AiDGX agent

arXiv:2502.15761v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) on extended reality (XR) devices has great potential to advance the field of human-AI interacti

Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding

SafetyDGX agent

arXiv:2602.02977v2 Announce Type: replace-cross Abstract: Vision-language models such as CLIP often struggle to faithfully understand long, detail-rich captions, relying on dominant scene cues while o

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models

ResearchDGX agent

arXiv:2605.13010v1 Announce Type: cross Abstract: We study image inpainting with generative diffusion models. Existing methods typically either train dedicated task-specific models, or adapt a pretrai

Amplification to Synthesis: A Comparative Analysis of Cognitive Operations Before and After Generative AI

ResearchDGX agent

arXiv:2605.13785v1 Announce Type: cross Abstract: Cognitive operations are a rising concern in the geopolitical sphere, a quiet yet rigorous fight for public perception and decision making. While such

An Agentic AI Framework with Large Language Models and Chain-of-Thought for UAV-Assisted Logistics Scheduling with Mobile Edge Computing

Local AiDGX agent

arXiv:2605.13221v1 Announce Type: new Abstract: In cloud manufacturing, unmanned aerial vehicles (UAVs) can support both product collection and mobile edge computing (MEC). This joint operation forms

An Agentic LLM-Based Framework for Population-Scale Mental Health Screening

AgentsDGX agent

arXiv:2605.13046v1 Announce Type: new Abstract: Mental health disorders affect millions worldwide, and healthcare systems are increasingly overwhelmed by the volume of clinical data generated from ele

An Efficient Insect-inspired Approach for Visual Point-goal Navigation

Model ReleasesDGX agent

arXiv:2601.16806v3 Announce Type: replace Abstract: In this work we develop a novel insect-inspired model for visual point-goal navigation. This combines abstracted models of two insect brain structur

Anatomy-Slot: Unsupervised Anatomical Factorization for Homologous Bilateral Reasoning in Retinal Diagnosis

ResearchDGX agent

arXiv:2605.12929v1 Announce Type: cross Abstract: Retinal diagnosis is inherently bilateral: clinicians compare homologous structures across eyes (e.g., optic disc asymmetry), yet most deep models ope

AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation

SafetyDGX agent

arXiv:2605.13724v1 Announce Type: cross Abstract: Few-step video generation has been significantly advanced by consistency distillation. However, the performance of consistency-distilled models often

ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive Margin

ResearchDGX agent

arXiv:2605.13517v1 Announce Type: cross Abstract: Vector Quantized Variational Autoencoder (VQ-VAE) has become a fundamental framework for learning discrete representations in image modeling. However,

Are Compact Rationales Free? Measuring Tile Selection Headroom in Frozen WSI-MIL

ResearchDGX agent

arXiv:2605.12575v1 Announce Type: cross Abstract: Whole-slide image (WSI) multiple instance learning (MIL) classifiers can achieve strong slide-level AUC while leaving the full-bag prediction opaque.

AssemblyBench: Physics-Aware Assembly of Complex Industrial Objects

ResearchDGX agent

arXiv:2605.12845v1 Announce Type: cross Abstract: Assembling objects from parts requires understanding multimodal instructions, linking them to 3D components, and predicting physically plausible 6-DoF

Assessing the Creativity of Large Language Models: Testing, Limits, and New Frontiers

ResearchDGX agent

arXiv:2605.13450v1 Announce Type: new Abstract: Measuring the creativity of large language models (LLMs) is essential for designing methods that can improve creativity and for enhancing our scientific

AttenA+: Rectifying Action Inequality in Robotic Foundation Models

Model ReleasesDGX agent

arXiv:2605.13548v1 Announce Type: cross Abstract: Existing robotic foundation models, while powerful, are predicated on an implicit assumption of temporal homogeneity: treating all actions as equally

Auditing Sybil: Explaining Deep Lung Cancer Risk Prediction Through Generative Interventional Attributions

SafetyDGX agent

arXiv:2602.02560v2 Announce Type: replace-cross Abstract: Lung cancer remains the leading cause of cancer mortality, driving the development of automated screening tools to alleviate radiologist workl

AuraMask: An Extensible Pipeline for Developing Aesthetic Anti-Facial Recognition Image Filters

ResearchDGX agent

arXiv:2605.12937v1 Announce Type: cross Abstract: Anti-facial recognition (AFR) image filters alter images in ways that are subtle to people but blinding to computer vision. Yet, despite widespread in

Automated alignment is harder than you think

SafetyDGX agent

arXiv:2605.06390v2 Announce Type: replace Abstract: A leading proposal for aligning artificial superintelligence (ASI) is to use AI agents to automate an increasing fraction of alignment research as c

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

SafetyDGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

Bayesian Model Merging

Model ReleasesDGX agent

arXiv:2605.12843v1 Announce Type: cross Abstract: Model merging aims to combine multiple task-specific expert models into a single model without joint retraining, offering a practical alternative to m

BEAVER: An Enterprise Benchmark for Text-to-SQL

Model ReleasesDGX agent

arXiv:2409.02038v3 Announce Type: replace-cross Abstract: Existing text-to-SQL benchmarks have largely been constructed from public databases with well-structured schemas and simplistic question-SQL p

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

SafetyDGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

Beyond Anthropomorphism: Exploring the Roles of Perceived Non-humanity and Structural Similarity in Deep Self-Disclosure Toward Generative AI

SafetyDGX agent

arXiv:2605.13574v1 Announce Type: cross Abstract: This study investigates deep self-disclosure toward generative AI by examining perceived non-humanity and structural similarity as psychological facto

Beyond Cooperative Simulators: Generating Realistic User Personas for Robust Evaluation of LLM Agents

ResearchDGX agent

arXiv:2605.12894v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly deployed in settings where they interact with a wide variety of people, including users who are uncle

Beyond Individual Mimicry: Constructing Human-Like Social network with Graph-Augmented LLM Agents

Local AiDGX agent

arXiv:2605.12512v1 Announce Type: cross Abstract: Driven by large language models (LLMs), social bot can autonomously engage in local interactions, whose human-like behaviors enable them to evade soci

Beyond Perplexity: A Geometric and Spectral Study of Low-Rank Pre-Training

ResearchDGX agent

arXiv:2605.13652v1 Announce Type: cross Abstract: Pre-training large language models is dominated by the memory cost of storing full-rank weights, gradients, and optimizer states. Low-rank pre-trainin

Block-wise Adaptive Caching for Accelerating Diffusion Policy

SafetyDGX agent

arXiv:2506.13456v2 Announce Type: replace Abstract: Diffusion Policy has demonstrated strong visuomotor modeling capabilities, but its high computational cost renders it impractical for real-time robo

BoostTaxo: Zero-Shot Taxonomy Induction via Boosting-Style Agentic Reasoning and Constraint-Aware Calibration

Model ReleasesDGX agent

arXiv:2605.12520v1 Announce Type: cross Abstract: Taxonomy induction is crucial for organizing concepts into explicit and interpretable semantic hierarchies. While existing methods have achieved promi

Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.13054v1 Announce Type: cross Abstract: Cross-domain offline reinforcement learning aims to adapt a policy from a source domain to a target domain using only pre-collected datasets, where en

Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models

ResearchDGX agent

arXiv:2605.12517v1 Announce Type: cross Abstract: Vision-language models (VLMs) are often deployed on text-only inputs, although they are trained with images. We find that removing the vision modality

CADDesigner: Conceptual CAD Model Generation with a General-Purpose Agent

AgentsDGX agent

arXiv:2508.01031v5 Announce Type: replace Abstract: Computer-Aided Design (CAD) is widely used for conceptual design and parametric 3D modeling, but typically requires a high level of expertise from d

Can LLM Agents Simulate Dynamic Networks? A Case Study on Email Networks with Phishing Synthesis

AgentsDGX agent

arXiv:2605.12507v1 Announce Type: cross Abstract: While Large Language Model (LLM) multi-agent systems (MAS) offer a transformative approach to simulating human behavior in complex systems, it remains

CANTANTE: Optimizing Agentic Systems via Contrastive Credit Attribution

Local AiDGX agent

arXiv:2605.13295v1 Announce Type: cross Abstract: LLM-based multi-agent systems have demonstrated strong performance across complex real-world tasks, such as software engineering, predictive modeling,

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

SafetyDGX agent

arXiv:2605.13646v1 Announce Type: cross Abstract: End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recentl

CHAL: Council of Hierarchical Agentic Language

AgentsDGX agent

arXiv:2605.12718v1 Announce Type: new Abstract: Multi-agent debate has emerged as a promising approach for improving LLM reasoning on ground-truth tasks, yet current methodologies face certain structu

ChannelKAN: Multi-Scale Dual-Domain Channel Prediction via Hybrid CNN-KAN Architecture

Local AiDGX agent

arXiv:2605.12553v1 Announce Type: cross Abstract: Accurate channel state information (CSI) prediction is essential for improving the reliability and spectral efficiency of massive MIMO-OFDM systems in

Characteristic Root Analysis and Regularization for Linear Time Series Forecasting

TutorialsDGX agent

arXiv:2509.23597v5 Announce Type: replace-cross Abstract: Time series forecasting remains a critical challenge across numerous domains, yet the effectiveness of complex models often varies unpredictab

ChatSR: Multimodal Large Language Models for Scientific Formula Discovery

SafetyDGX agent

arXiv:2406.05410v3 Announce Type: replace Abstract: Current multimodal large language models (MLLMs) are mainly focused on the understanding and processing of perceptual modalities such as images and

Children's English Reading Story Generation via Supervised Fine-Tuning of Compact LLMs with Controllable Difficulty and Safety

Model ReleasesDGX agent

arXiv:2605.13709v1 Announce Type: cross Abstract: Large Language Models (LLMs) are widely applied in educational practices, such as for generating children's stories. However, the generated stories ar

ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation

Model ReleasesDGX agent

arXiv:2605.12857v1 Announce Type: cross Abstract: Existing API-based agentic systems for RTL code generation are fundamentally misaligned with industrial practice: they assume a golden testbench is av

CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models

ResearchDGX agent

arXiv:2605.13178v1 Announce Type: cross Abstract: In large vision-language models, visual tokens typically constitute the majority of input tokens, leading to substantial computational overhead. To ad

CodeClash: Benchmarking Goal-Oriented Software Engineering

Model ReleasesDGX agent

arXiv:2511.00839v2 Announce Type: replace-cross Abstract: Current benchmarks for coding evaluate language models (LMs) on concrete, well-specified tasks such as fixing specific bugs or writing targete

CoGE: Sim-to-Real Online Geometric Estimation for Monocular Colonoscopy

Local AiDGX agent

arXiv:2605.13038v1 Announce Type: cross Abstract: Geometric estimation including depth estimation and scene reconstruction is a crucial technique for colonoscopy which can provide surgeons with 3D spa

Cognifold: Always-On Proactive Memory via Cognitive Folding

AgentsDGX agent

arXiv:2605.13438v1 Announce Type: new Abstract: Existing agent memory remains predominantly reactive and retrieval-based, lacking the capacity to autonomously organize experience into persistent cogni

Compact Latent Manifold Translation: A Parameter-Efficient Foundation Model for Cross-Modal and Cross-Frequency Physiological Signal Synthesis

Model ReleasesDGX agent

arXiv:2605.13248v1 Announce Type: cross Abstract: The analysis of physiological time series, such as electrocardiograms (ECG) and photoplethysmograms (PPG), is persistently hindered by modality and fr

Constitutional Governance in Metric Spaces

ResearchDGX agent

arXiv:2605.13362v1 Announce Type: cross Abstract: Computational social choice and algorithmic decision theory offer rich aggregation theory but no end-to-end, polynomial-time process for egalitarian s

Context Matters: Auditing Gender Bias in T2I Generation through Risk-Tiered Use-Case Profiles

SafetyDGX agent

arXiv:2605.13113v1 Announce Type: cross Abstract: Text-to-image (T2I) generative models are increasingly used to produce content for education, media, and public-facing communication, and are starting

Context Training with Active Information Seeking

ResearchDGX agent

arXiv:2605.13050v1 Announce Type: cross Abstract: Most existing large language models (LLMs) are expensive to adapt after deployment, especially when a task requires newly produced information or nich

Continual Learning with Multilingual Foundation Model

ResearchDGX agent

arXiv:2605.13415v1 Announce Type: cross Abstract: This paper presents a multi-stage framework for detecting reclaimed slurs in multilingual social media discourse. It addresses the challenge of identi

Controllable Quantum Memory Capacity in Quantum Reservoir Networks with Tunable partial-SWAPs

Model ReleasesDGX agent

arXiv:2605.12713v1 Announce Type: cross Abstract: In the field of quantum reservoir computing (QRC), many different computational models and architectures have been proposed. From these models, we ide

Controlling Logical Collapse in LLMs via Algebraic Ontology Projection over F2

Model ReleasesDGX agent

arXiv:2605.12968v1 Announce Type: cross Abstract: Do large language models internally encode ontological relations in a formally verifiable algebraic structure? We introduce Algebraic Ontology Project

Coordinating Multiple Conditions for Trajectory-Controlled Human Motion Generation

ResearchDGX agent

arXiv:2605.13729v1 Announce Type: cross Abstract: Trajectory-controlled human motion generation aims to synthesize realistic human motions conditioned on both textual descriptions and spatial trajecto

CoRe-Gen: Robust Spectrum-to-Structure Generation under Imperfect Fingerprint Conditions

Model ReleasesDGX agent

arXiv:2605.12980v1 Announce Type: cross Abstract: Molecular structure elucidation from tandem mass spectra (MS/MS) remains challenging, particularly for de novo generation beyond database coverage. A

Correct Answers from Sound Reasoning: Verifiable Process Supervision for Language Models

ResearchDGX agent

arXiv:2605.12519v1 Announce Type: cross Abstract: Training language models to produce both correct answers and sound reasoning remains an open challenge. Reinforcement learning with verifiable rewards

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces

TutorialsDGX agent

arXiv:2605.12809v1 Announce Type: cross Abstract: A critical step for reliable large language models (LLMs) use in healthcare is to attribute predictions to their training data, akin to a medical case

← Previous
1…256257258259260…358
Next →