AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
10 Apr 2026

ClawLess: A Security Model of AI Agents

AgentsDGX agent

arXiv:2604.06284v1 Announce Type: cross Abstract: Autonomous AI agents powered by Large Language Models can reason, plan, and execute complex tasks, but their ability to autonomously retrieve informat

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

Model ReleasesDGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

CNN-based Surface Temperature Forecasts with Ensemble Numerical Weather Prediction

SafetyDGX agent

arXiv:2507.18937v3 Announce Type: replace-cross Abstract: Due to limited computational resources, medium-range temperature forecasts typically rely on low-resolution numerical weather prediction (NWP)


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Code Sharing In Prediction Model Research: A Scoping Review

ResearchDGX agent

arXiv:2604.06212v1 Announce Type: cross Abstract: Analytical code is essential for reproducing diagnostic and prognostic prediction model research, yet code availability in the published literature re

Commander-GPT: Dividing and Routing for Multimodal Sarcasm Detection

Model ReleasesDGX agent

arXiv:2506.19420v2 Announce Type: replace Abstract: Multimodal sarcasm understanding is a high-order cognitive task. Although large language models (LLMs) have shown impressive performance on many dow

Compressible Softmax-Attended Language under Incompressible Attention

ResearchDGX agent

arXiv:2604.04384v2 Announce Type: replace-cross Abstract: Softmax attention defines an interaction through d_h head dimensions, but not all dimensions carry equal weight once real text passes throug

Computer Environments Elicit General Agentic Intelligence in LLMs

AgentsDGX agent

arXiv:2601.16206v3 Announce Type: replace-cross Abstract: Agentic intelligence in large language models (LLMs) requires not only model intrinsic capabilities but also interactions with external enviro

Concentrated siting of AI data centers drives regional power-system stress under rising global compute demand

Local AiDGX agent

arXiv:2604.06198v1 Announce Type: cross Abstract: The rapid rise of generative artificial intelligence (AI) is driving unprecedented growth in global computational demand, placing increasing pressure

ConceptTracer: Interactive Analysis of Concept Saliency and Selectivity in Neural Representations

ResearchDGX agent

arXiv:2604.07019v1 Announce Type: cross Abstract: Neural networks deliver impressive predictive performance across a variety of tasks, but they are often opaque in their decision-making processes. Des

ConfusionPrompt: Practical Private Inference for Online Large Language Models

Local AiDGX agent

arXiv:2401.00870v5 Announce Type: replace-cross Abstract: State-of-the-art large language models (LLMs) are typically deployed as online services, requiring users to transmit detailed prompts to cloud

Consistency-Guided Decoding with Proof-Driven Disambiguation for Three-Way Logical Question Answering

Model ReleasesDGX agent

arXiv:2604.06196v1 Announce Type: cross Abstract: Three-way logical question answering (QA) assigns True/False/Unknown to a hypothesis H given a premise set S. While modern large language models

Continual Visual Anomaly Detection on the Edge: Benchmark and Efficient Solutions

Model ReleasesDGX agent

arXiv:2604.06435v1 Announce Type: cross Abstract: Visual Anomaly Detection (VAD) is a critical task for many applications including industrial inspection and healthcare. While VAD has been extensively

Contrastive Decoding Mitigates Score Range Bias in LLM-as-a-Judge

SafetyDGX agent

arXiv:2510.18196v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are commonly used as evaluators in various applications, but the reliability of the outcomes remains a challenge.

ConvoLearn: A Dataset for Fine-Tuning Dialogic AI Tutors

Model ReleasesDGX agent

arXiv:2601.08950v3 Announce Type: replace Abstract: Despite their growing adoption in education, LLMs remain misaligned with the core principle of effective tutoring: the dialogic construction of know

Countering the Over-Reliance Trap: Mitigating Object Hallucination for LVLMs via a Self-Validation Framework

ResearchDGX agent

arXiv:2601.22451v2 Announce Type: replace-cross Abstract: Despite progress in Large Vision Language Models (LVLMs), object hallucination remains a critical issue in image captioning task, where models

Cross-Lingual Transfer and Parameter-Efficient Adaptation in the Turkic Language Family: A Theoretical Framework for Low-Resource Language Models

Model ReleasesDGX agent

arXiv:2604.06202v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed natural language processing, yet their capabilities remain uneven across languages. Most multilingual mo

CSA-Graphs: A Privacy-Preserving Structural Dataset for Child Sexual Abuse Research

SafetyDGX agent

arXiv:2604.07132v1 Announce Type: cross Abstract: Child Sexual Abuse Imagery (CSAI) classification is an important yet challenging problem for computer vision research due to the strict legal and ethi

CubeGraph: Efficient Retrieval-Augmented Generation for Spatial and Temporal Data

ApplicationsDGX agent

arXiv:2604.06616v1 Announce Type: cross Abstract: Hybrid queries combining high-dimensional vector similarity search with spatio-temporal filters are increasingly critical for modern retrieval-augment

Daily and Weekly Periodicity in Large Language Model Performance and Its Implications for Research

ResearchDGX agent

arXiv:2602.15889v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used in research as both tools and objects of study. Much of this work assumes that LLM performa

Database Querying under Missing Values Governed by Missingness Mechanisms

ResearchDGX agent

arXiv:2604.06520v1 Announce Type: cross Abstract: We address the problems of giving a semantics to- and doing query answering (QA) on a relational database (RDB) that has missing values (MVs). The cau

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation

ResearchDGX agent

arXiv:2511.19365v2 Announce Type: replace-cross Abstract: Pixel diffusion aims to generate images directly in pixel space in an end-to-end fashion. This approach avoids the limitations of VAE in the t

Depression Detection at the Point of Care: Automated Analysis of Linguistic Signals from Routine Primary Care Encounters

ResearchDGX agent

arXiv:2604.06193v1 Announce Type: cross Abstract: Depression is underdiagnosed in primary care, yet timely identification remains critical. Recorded clinical encounters, increasingly common with digit

Designing Safe and Accountable GenAI as a Learning Companion with Women Banned from Formal Education

SafetyDGX agent

arXiv:2604.07253v1 Announce Type: cross Abstract: In gender-restrictive and surveilled contexts, where access to formal education may be restricted for women, pursuing education involves safety and pr

Development of ML model for triboelectric nanogenerator based sign language detection system

ResearchDGX agent

arXiv:2604.06220v1 Announce Type: cross Abstract: Sign language recognition (SLR) is vital for bridging communication gaps between deaf and hearing communities. Vision-based approaches suffer from occ

Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment

Model ReleasesDGX agent

arXiv:2601.08258v3 Announce Type: replace Abstract: Large language models increasingly fail in a way that scalar accuracy cannot diagnose: they produce a sound reasoning trace and then abandon it unde

DietDelta: A Vision-Language Approach for Dietary Assessment via Before-and-After Images

ResearchDGX agent

arXiv:2604.06352v1 Announce Type: cross Abstract: Accurate dietary assessment is critical for precision nutrition, yet most image-based methods rely on a single pre-consumption image and provide only

DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models

TutorialsDGX agent

arXiv:2306.14685v5 Announce Type: replace-cross Abstract: We demonstrate that pre-trained text-to-image diffusion models, despite being trained on raster images, possess a remarkable capacity to guide

Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji Embeddings

Model ReleasesDGX agent

arXiv:2604.06863v1 Announce Type: cross Abstract: Skin-toned emojis are crucial for fostering personal identity and social inclusion in online communication. As AI models, particularly Large Language

Discrete Flow Matching Policy Optimization

SafetyDGX agent

arXiv:2604.06491v1 Announce Type: cross Abstract: We introduce Discrete flow Matching policy Optimization (DoMinO), a unified framework for Reinforcement Learning (RL) fine-tuning Discrete Flow Matchi

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

Model ReleasesDGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

Distributed Interpretability and Control for Large Language Models

Model ReleasesDGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook

SafetyDGX agent

arXiv:2604.06210v2 Announce Type: cross Abstract: As LLMs are globally deployed, aligning their cultural value orientations is critical for safety and user engagement. However, existing benchmarks fac

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

ResearchDGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

Domain-Contextualized Inference: A Computable Graph Architecture for Explicit-Domain Reasoning

Model ReleasesDGX agent

arXiv:2604.04344v2 Announce Type: replace Abstract: We establish a computation-substrate-agnostic inference architecture in which domain is an explicit first-class computational parameter. This produc

'Don't Be Afraid, Just Learn': Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI

TutorialsDGX agent

arXiv:2604.06342v1 Announce Type: cross Abstract: Although tension between university curricula and industry expectations has existed in some form for decades, the rapid integration of generative AI (

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

Model ReleasesDGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

Model ReleasesDGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

Dynamic Context Evolution for Scalable Synthetic Data Generation

Model ReleasesDGX agent

arXiv:2604.07147v1 Announce Type: cross Abstract: Large language models produce repetitive output when prompted independently across many batches, a phenomenon we term cross-batch mode collapse: the p

Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees

ResearchDGX agent

arXiv:2604.06515v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) allows scaling of language and vision models efficiently by activating only a small subset of experts per input. While

EmoMAS: Emotion-Aware Multi-Agent System for High-Stakes Edge-Deployable Negotiation with Bayesian Orchestration

Local AiDGX agent

arXiv:2604.07003v1 Announce Type: new Abstract: Large language models (LLMs) has been widely used for automated negotiation, but their high computational cost and privacy risks limit deployment in pri

Energy-based Tissue Manifolds for Longitudinal Multiparametric MRI Analysis

TutorialsDGX agent

arXiv:2604.07180v1 Announce Type: cross Abstract: We propose a geometric framework for longitudinal multi-parametric MRI analysis based on patient-specific energy modelling in sequence space. Rather t

Energy Saving for Cell-Free Massive MIMO Networks: A Multi-Agent Deep Reinforcement Learning Approach

AgentsDGX agent

arXiv:2604.07133v1 Announce Type: cross Abstract: This paper focuses on energy savings in downlink operation of cell-free massive MIMO (CF mMIMO) networks under dynamic traffic conditions. We propose

Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Models

ApplicationsDGX agent

arXiv:2604.06826v1 Announce Type: cross Abstract: Environmental, Social, and Governance (ESG) considerations are increasingly integral to assessing corporate performance, reputation, and long-term sus

Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction

ResearchDGX agent

arXiv:2604.07320v1 Announce Type: cross Abstract: Low-resource languages pose a challenge for machine translation with large language models (LLMs), which require large amounts of training data. One p

Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios

Model ReleasesDGX agent

arXiv:2604.06742v1 Announce Type: cross Abstract: Large Language Models (LLMs) are driving a shift towards intent-driven development, where agents build complete software from scratch. However, existi

Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development

Model ReleasesDGX agent

arXiv:2604.06793v1 Announce Type: cross Abstract: Software documentation is crucial for repository comprehension. While Large Language Models (LLMs) advance documentation generation from code snippets

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

Model ReleasesDGX agent

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

EviSnap: Faithful Evidence-Cited Explanations for Cold-Start Cross-Domain Recommendation

ResearchDGX agent

arXiv:2604.06172v1 Announce Type: cross Abstract: Cold-start cross-domain recommender (CDR) systems predict a user's preferences in a target domain using only their source-domain behavior, yet existin

Explaining Neural Networks in Preference Learning: a Post-hoc Inductive Logic Programming Approach

Local AiDGX agent

arXiv:2604.06838v1 Announce Type: new Abstract: In this paper, we propose using Learning from Answer Sets to approximate black-box models, such as Neural Networks (NN), in the specific case of learnin

Exploring Natural Language-Based Strategies for Efficient Number Learning in Children through Reinforcement Learning

ApplicationsDGX agent

arXiv:2410.08334v2 Announce Type: replace-cross Abstract: In this paper, we build a reinforcement learning framework to study how children compose numbers using base-ten blocks. Studying numerical cog

Extracting Breast Cancer Phenotypes from Clinical Notes: Comparing LLMs with Classical Ontology Methods

ResearchDGX agent

arXiv:2604.06208v1 Announce Type: cross Abstract: A significant amount of data held in Oncology Electronic Medical Records (EMRs) is contained in unstructured provider notes -- including but not limit

Faithful-First Reasoning, Planning, and Acting for Multimodal LLMs

ResearchDGX agent

arXiv:2511.08409v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) frequently suffer from unfaithfulness, generating reasoning chains that drift from visual evidence or contr

FBS: Modeling Native Parallel Reading inside a Transformer

ResearchDGX agent

arXiv:2601.21708v2 Announce Type: replace Abstract: Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing accelerat

FedDAP: Domain-Aware Prototype Learning for Federated Learning under Domain Shift

Local AiDGX agent

arXiv:2604.06795v1 Announce Type: cross Abstract: Federated Learning (FL) enables decentralized model training across multiple clients without exposing private data, making it ideal for privacy-sensit

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

Fine-grained Approaches for Confidence Calibration of LLMs in Automated Code Revision

ResearchDGX agent

arXiv:2604.06723v1 Announce Type: cross Abstract: In today's AI-assisted software engineering landscape, developers increasingly depend on LLMs that are highly capable, yet inherently imperfect. The t

FLeX: Fourier-based Low-rank EXpansion for multilingual transfer

Model ReleasesDGX agent

arXiv:2604.06253v1 Announce Type: cross Abstract: Cross-lingual code generation is critical in enterprise environments where multiple programming languages coexist. However, fine-tuning large language

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

Model ReleasesDGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

FlowExtract: Procedural Knowledge Extraction from Maintenance Flowcharts

ApplicationsDGX agent

arXiv:2604.06770v1 Announce Type: cross Abstract: Maintenance procedures in manufacturing facilities are often documented as flowcharts in static PDFs or scanned images. They encode procedural knowled

← Previous
1…344345346347348…350
Next →