AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
10 Apr 2026

Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios

Model ReleasesDGX agent

arXiv:2604.06742v1 Announce Type: cross Abstract: Large Language Models (LLMs) are driving a shift towards intent-driven development, where agents build complete software from scratch. However, existi

Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development

Model ReleasesDGX agent

arXiv:2604.06793v1 Announce Type: cross Abstract: Software documentation is crucial for repository comprehension. While Large Language Models (LLMs) advance documentation generation from code snippets

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

EviSnap: Faithful Evidence-Cited Explanations for Cold-Start Cross-Domain Recommendation

ResearchDGX agent

arXiv:2604.06172v1 Announce Type: cross Abstract: Cold-start cross-domain recommender (CDR) systems predict a user's preferences in a target domain using only their source-domain behavior, yet existin

Explaining Neural Networks in Preference Learning: a Post-hoc Inductive Logic Programming Approach

Local AiDGX agent

arXiv:2604.06838v1 Announce Type: new Abstract: In this paper, we propose using Learning from Answer Sets to approximate black-box models, such as Neural Networks (NN), in the specific case of learnin

Exploring Natural Language-Based Strategies for Efficient Number Learning in Children through Reinforcement Learning

ApplicationsDGX agent

arXiv:2410.08334v2 Announce Type: replace-cross Abstract: In this paper, we build a reinforcement learning framework to study how children compose numbers using base-ten blocks. Studying numerical cog

Extracting Breast Cancer Phenotypes from Clinical Notes: Comparing LLMs with Classical Ontology Methods

ResearchDGX agent

arXiv:2604.06208v1 Announce Type: cross Abstract: A significant amount of data held in Oncology Electronic Medical Records (EMRs) is contained in unstructured provider notes -- including but not limit

Faithful-First Reasoning, Planning, and Acting for Multimodal LLMs

ResearchDGX agent

arXiv:2511.08409v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) frequently suffer from unfaithfulness, generating reasoning chains that drift from visual evidence or contr

FBS: Modeling Native Parallel Reading inside a Transformer

ResearchDGX agent

arXiv:2601.21708v2 Announce Type: replace Abstract: Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing accelerat

FedDAP: Domain-Aware Prototype Learning for Federated Learning under Domain Shift

Local AiDGX agent

arXiv:2604.06795v1 Announce Type: cross Abstract: Federated Learning (FL) enables decentralized model training across multiple clients without exposing private data, making it ideal for privacy-sensit

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

Fine-grained Approaches for Confidence Calibration of LLMs in Automated Code Revision

ResearchDGX agent

arXiv:2604.06723v1 Announce Type: cross Abstract: In today's AI-assisted software engineering landscape, developers increasingly depend on LLMs that are highly capable, yet inherently imperfect. The t

FLeX: Fourier-based Low-rank EXpansion for multilingual transfer

Model ReleasesDGX agent

arXiv:2604.06253v1 Announce Type: cross Abstract: Cross-lingual code generation is critical in enterprise environments where multiple programming languages coexist. However, fine-tuning large language

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

Model ReleasesDGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

FlowExtract: Procedural Knowledge Extraction from Maintenance Flowcharts

ApplicationsDGX agent

arXiv:2604.06770v1 Announce Type: cross Abstract: Maintenance procedures in manufacturing facilities are often documented as flowcharts in static PDFs or scanned images. They encode procedural knowled

FMI@SU ToxHabits: Evaluating LLMs Performance on Toxic Habit Extraction in Spanish Clinical Texts

Model ReleasesDGX agent

arXiv:2604.06403v1 Announce Type: cross Abstract: The paper presents an approach for the recognition of toxic habits named entities in Spanish clinical texts. The approach was developed for the ToxHab

FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling

SafetyDGX agent

arXiv:2604.06916v1 Announce Type: cross Abstract: Reinforcement-Learning-based post-training has recently emerged as a promising paradigm for aligning text-to-image diffusion models with human prefere

Frailty Estimation in Elderly Oncology Patients Using Multimodal Wearable Data and Multi-Instance Learning

ApplicationsDGX agent

arXiv:2604.06985v1 Announce Type: cross Abstract: Frailty and functional decline strongly influence treatment tolerance and outcomes in older patients with cancer, yet assessment is typically limited

From experimentation to engagement: on the paradox of participatory AI and power in contexts of forced displacement and humanitarian crises

SafetyDGX agent

arXiv:2604.06219v1 Announce Type: cross Abstract: Across the Global North, calls for participatory artificial intelligence (AI) to improve the responsible, safe, and ethical use of AI have increased,

From Exploration to Revelation: Detecting Dark Patterns in Mobile Apps

ResearchDGX agent

arXiv:2411.18084v2 Announce Type: replace-cross Abstract: Mobile apps are essential in daily life but frequently employ deceptive patterns, such as visual emphasis or linguistic nudging, to manipulate

From Exposure to Internalization: Dual-Stream Calibration for In-context Clinical Reasoning

ResearchDGX agent

arXiv:2604.06262v1 Announce Type: cross Abstract: Contextual clinical reasoning demands robust inference grounded in complex, heterogeneous clinical records. While state-of-the-art fine-tuning, in-con

From Load Tests to Live Streams: Graph Embedding-Based Anomaly Detection in Microservice Architectures

ResearchDGX agent

arXiv:2604.06448v1 Announce Type: cross Abstract: Prime Video regularly conducts load tests to simulate the viewer traffic spikes seen during live events such as Thursday Night Football as well as vid

Front-End Ethics for Sensor-Fused Health Conversational Agents: An Ethical Design Space for Biometrics

SafetyDGX agent

arXiv:2604.06203v1 Announce Type: cross Abstract: The integration of continuous data from built-in sensors and Large Language Models (LLMs) has fueled a surge of 'Sensor-Fused LLM agents' for personal

Full State-Space Visualisation of the 8-Puzzle: Feasibility, Design, and Educational Use

HardwareDGX agent

arXiv:2604.06186v1 Announce Type: cross Abstract: Search algorithms are a foundational topic in artificial intelligence education, yet even simple domains can generate large state spaces that challeng

FVD: Inference-Time Alignment of Diffusion Models via Fleming-Viot Resampling

SafetyDGX agent

arXiv:2604.06779v1 Announce Type: new Abstract: We introduce Fleming-Viot Diffusion (FVD), an inference-time alignment method that resolves the diversity collapse commonly observed in Sequential Monte

Generating Attribution Reports for Manipulated Facial Images: A Dataset and Baseline

Model ReleasesDGX agent

arXiv:2412.19685v2 Announce Type: replace-cross Abstract: Existing facial forgery detection methods typically focus on binary classification or pixel-level localization, providing little semantic insi

Generative Phomosaic with Structure-Aligned and Personalized Diffusion

ResearchDGX agent

arXiv:2604.06989v1 Announce Type: cross Abstract: We present the first generative approach to photomosaic creation. Traditional photomosaic methods rely on a large number of tile images and color-base

Governance and Regulation of Artificial Intelligence in Developing Countries: A Case Study of Nigeria

Local AiDGX agent

arXiv:2604.06018v2 Announce Type: replace-cross Abstract: This study examines the perception of legal professionals on the governance of AI in developing countries, using Nigeria as a case study. The

Governing frontier general-purpose AI in the public sector: adaptive risk management and policy capacity under uncertainty through 2030

SafetyDGX agent

arXiv:2604.06215v1 Announce Type: cross Abstract: The governance of frontier general-purpose artificial intelligence has become a public-sector problem of institutional design, not merely a technical

GS-Surrogate: Deformable Gaussian Splatting for Parameter Space Exploration of Ensemble Simulations

Model ReleasesDGX agent

arXiv:2604.06358v1 Announce Type: cross Abstract: Exploring ensemble simulations is increasingly important across many scientific domains. However, supporting flexible post-hoc exploration remains cha

Hallucination as output-boundary misclassification: a composite abstention architecture for language models

ResearchDGX agent

arXiv:2604.06195v1 Announce Type: cross Abstract: Large language models often produce unsupported claims. We frame this as a misclassification error at the output boundary, where internally generated

Harf-Speech: A Clinically Aligned Framework for Arabic Phoneme-Level Speech Assessment

Model ReleasesDGX agent

arXiv:2604.06191v1 Announce Type: cross Abstract: Automated phoneme-level pronunciation assessment is vital for scalable speech therapy and language learning, yet validated tools for Arabic remain sca

Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization

SafetyDGX agent

arXiv:2604.06285v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have become essential for tasks such as image synthesis, captioning, and retrieval by aligning textual and visual inform

High-Precision Estimation of the State-Space Complexity of Shogi via the Monte Carlo Method

ApplicationsDGX agent

arXiv:2604.06189v1 Announce Type: new Abstract: Determining the state-space complexity of the game of Shogi (Japanese Chess) has been a challenging problem, with previous combinatorial estimates leavi

HingeMem: Boundary Guided Long-Term Memory with Query Adaptive Retrieval for Scalable Dialogues

Model ReleasesDGX agent

arXiv:2604.06845v1 Announce Type: cross Abstract: Long-term memory is critical for dialogue systems that support continuous, sustainable, and personalized interactions. However, existing methods rely

How Much LLM Does a Self-Revising Agent Actually Need?

AgentsDGX agent

arXiv:2604.07236v2 Announce Type: new Abstract: Recent LLM-based agents often place world modeling, planning, and reflection inside a single language model loop. This can produce capable behavior, but

How to Evaluate Speech Translation with Source-Aware Neural MT Metrics

SafetyDGX agent

arXiv:2511.03295v3 Announce Type: replace-cross Abstract: Automatic evaluation of ST systems is typically performed by comparing translation hypotheses with one or more reference translations. While e

HQF-Net: A Hybrid Quantum-Classical Multi-Scale Fusion Network for Remote Sensing Image Segmentation

Local AiDGX agent

arXiv:2604.06715v1 Announce Type: cross Abstract: Remote sensing semantic segmentation requires models that can jointly capture fine spatial details and high-level semantic context across complex scen

Hybrid ResNet-1D-BiGRU with Multi-Head Attention for Cyberattack Detection in Industrial IoT Environments

ResearchDGX agent

arXiv:2604.06481v1 Announce Type: cross Abstract: This study introduces a hybrid deep learning model for intrusion detection in Industrial IoT (IIoT) systems, combining ResNet-1D, BiGRU, and Multi-Hea

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

Model ReleasesDGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

Implantable Adaptive Cells: A Novel Enhancement for Pre-Trained U-Nets in Medical Image Segmentation

ResearchDGX agent

arXiv:2405.03420v2 Announce Type: cross Abstract: This paper introduces a novel approach to enhance the performance of pre-trained neural networks in medical image segmentation using gradient-based Ne

Improved Evidence Extraction and Metrics for Document Inconsistency Detection with LLMs

ResearchDGX agent

arXiv:2601.02627v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are becoming useful in many domains due to their impressive abilities that arise from large training datasets and

Improving Robustness In Sparse Autoencoders via Masked Regularization

ResearchDGX agent

arXiv:2604.06495v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are widely used in mechanistic interpretability to project LLM activations onto sparse latent spaces. However, sparsity alo

In-Context Decision Making for Optimizing Complex AutoML Pipelines

Model ReleasesDGX agent

arXiv:2508.13657v2 Announce Type: replace-cross Abstract: Combined Algorithm Selection and Hyperparameter Optimization (CASH) has been fundamental to traditional AutoML systems. However, with the adva

In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads

ResearchDGX agent

arXiv:2604.06356v1 Announce Type: cross Abstract: In-Context Learning (ICL) has been extensively studied in text-only Language Models, but remains largely unexplored in the speech domain. Here, we inv

Incentive-Aware Multi-Fidelity Optimization for Generative Advertising in Large Language Models

ResearchDGX agent

arXiv:2604.06263v1 Announce Type: cross Abstract: Generative advertising in large language model (LLM) responses requires optimizing sponsorship configurations under two strict constraints: the strate

Inference-Time Code Selection via Symbolic Equivalence Partitioning

ResearchDGX agent

arXiv:2604.06485v1 Announce Type: cross Abstract: 'Best-of-N' selection is a popular inference-time scaling method for code generation using Large Language Models (LLMs). However, to reliably identify

Information as Structural Alignment: A Dynamical Theory of Continual Learning

Model ReleasesDGX agent

arXiv:2604.07108v1 Announce Type: cross Abstract: Catastrophic forgetting is not an engineering failure. It is a mathematical consequence of storing knowledge as global parameter superposition. Existi

Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions

Model ReleasesDGX agent

arXiv:2602.09987v5 Announce Type: replace-cross Abstract: Influence functions are commonly used to attribute model behavior to training documents. We explore the reverse: crafting training data that i

Instance-Adaptive Parametrization for Amortized Variational Inference

Model ReleasesDGX agent

arXiv:2604.06796v1 Announce Type: cross Abstract: Latent variable models, including variational autoencoders (VAE), remain a central tool in modern deep generative modeling due to their scalability an

Invisible Influences: Investigating Implicit Intersectional Biases through Persona Engineering in Large Language Models

Model ReleasesDGX agent

arXiv:2604.06213v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at human-like language generation but often embed and amplify implicit, intersectional biases, especially under per

Invisible to Humans, Triggered by Agents: Stealthy Jailbreak Attacks on Mobile Vision-Language Agents

SafetyDGX agent

arXiv:2510.07809v4 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) empower autonomous mobile agents, yet their security under realistic mobile deployment constraints remain

JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency

Model ReleasesDGX agent

arXiv:2604.03044v2 Announce Type: replace-cross Abstract: We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performan

k-Maximum Inner Product Attention for Graph Transformers and the Expressive Power of GraphGPS

Model ReleasesDGX agent

arXiv:2604.03815v2 Announce Type: replace-cross Abstract: Graph transformers have shown promise in overcoming limitations of traditional graph neural networks, such as oversquashing and difficulties i

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

Model ReleasesDGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

KD-MARL: Resource-Aware Knowledge Distillation in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.06691v1 Announce Type: new Abstract: Real world deployment of multi agent reinforcement learning MARL systems is fundamentally constrained by limited compute memory and inference time. Whil

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis

Model ReleasesDGX agent

arXiv:2604.07034v1 Announce Type: cross Abstract: We present KITE, a training-free, keyframe-anchored, layout-grounded front-end that converts long robot-execution videos into compact, interpretable t

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

Model ReleasesDGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

Large Language Models for Outpatient Referral: Problem Definition, Benchmarking and Challenges

ApplicationsDGX agent

arXiv:2503.08292v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to outpatient referral tasks across healthcare systems. However, there is a lack of stan

LAsset: An LLM-assisted Security Asset Identification Framework for System-on-Chip (SoC) Verification

HardwareDGX agent

arXiv:2601.02624v2 Announce Type: replace-cross Abstract: The growing complexity of modern system-on-chip (SoC) and IP designs is making security assurance difficult day by day. One of the fundamental

← Previous
1…349350351352353354
Next →