AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
15 Apr 2026

DeepTest Tool Competition 2026: Benchmarking an LLM-Based Automotive Assistant

ResearchDGX agent

arXiv:2604.12615v1 Announce Type: new Abstract: This report summarizes the results of the first edition of the Large Language Model (LLM) Testing competition, held as part of the DeepTest workshop at

Designing Reliable LLM-Assisted Rubric Scoring for Constructed Responses: Evidence from Physics Exams

SafetyDGX agent

arXiv:2604.12227v1 Announce Type: new Abstract: Student responses in STEM assessments are often handwritten and combine symbolic expressions, calculations, and diagrams, creating substantial variation

Detecting and refurbishing ground truth errors during training of deep learning-based echocardiography segmentation models

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12832v1 Announce Type: cross Abstract: Deep learning-based medical image segmentation typically relies on ground truth (GT) labels obtained through manual annotation, but these can be prone

Development, Evaluation, and Deployment of a Multi-Agent System for Thoracic Tumor Board

AgentsDGX agent

arXiv:2604.12161v1 Announce Type: new Abstract: Tumor boards are multidisciplinary conferences dedicated to producing actionable patient care recommendations with live review of primary radiology and

Disposition Distillation at Small Scale: A Three-Arc Negative Result

Model ReleasesDGX agent

arXiv:2604.11867v1 Announce Type: cross Abstract: We set out to train behavioral dispositions (self-verification, uncertainty acknowledgment, feedback integration) into small language models (0.6B to

Distorted or Fabricated? A Survey on Hallucination in Video LLMs

ResearchDGX agent

arXiv:2604.12944v1 Announce Type: cross Abstract: Despite significant progress in video-language modeling, hallucinations remain a persistent challenge in Video Large Language Models (Vid-LLMs), refer

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding

SafetyDGX agent

arXiv:2604.12812v1 Announce Type: new Abstract: Existing Multimodal Large Language Models (MLLMs) suffer from significant performance degradation on the long document understanding task as document le

Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models

SafetyDGX agent

arXiv:2511.00710v4 Announce Type: replace Abstract: Recent studies posit that Reinforcement Learning with Verifiable Rewards (RLVR) primarily amplifies behaviors inherent to the pre-training distribut

Domain-Specific Latent Representations Improve the Fidelity of Diffusion-Based Medical Image Super-Resolution

Local AiDGX agent

arXiv:2604.12152v1 Announce Type: cross Abstract: Latent diffusion models for medical image super-resolution universally inherit variational autoencoders designed for natural photographs. We show that

DoseRAD2026 Challenge dataset: AI accelerated photon and proton dose calculation for radiotherapy

Model ReleasesDGX agent

arXiv:2604.12778v1 Announce Type: cross Abstract: Purpose: Accurate dose calculation is essential in radiotherapy for precise tumor irradiation while sparing healthy tissue. With the growing adoption

Drawing on Memory: Dual-Trace Encoding Improves Cross-Session Recall in LLM Agents

Model ReleasesDGX agent

arXiv:2604.12948v1 Announce Type: new Abstract: LLM agents with persistent memory store information as flat factual records, providing little context for temporal reasoning, change tracking, or cross-

DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems

SafetyDGX agent

arXiv:2509.19695v3 Announce Type: replace-cross Abstract: Task oriented dialog systems often rely on static exploration strategies that do not adapt to dynamic dialog contexts, leading to inefficient

Efficiency of Proportional Mechanisms in Online Auto-Bidding Advertising

ResearchDGX agent

arXiv:2604.12799v1 Announce Type: cross Abstract: The rise of automated bidding strategies in online advertising presents new challenges in designing and analyzing efficient auction mechanisms. In thi

Efficient Adversarial Training via Criticality-Aware Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.12780v1 Announce Type: cross Abstract: Vision Transformer (ViT) models have achieved remarkable performance across various vision tasks, with scalability being a key advantage when applied

Efficient and Scalable Granular-ball Graph Coarsening Method for Large-scale Graph Node Classification

ResearchDGX agent

arXiv:2603.29148v2 Announce Type: replace-cross Abstract: Graph Convolutional Network (GCN) is a model that can effectively handle graph data tasks and has been successfully applied. However, for larg

Efficient Semantic Image Communication for Traffic Monitoring at the Edge

ResearchDGX agent

arXiv:2604.12622v1 Announce Type: cross Abstract: Many visual monitoring systems operate under strict communication constraints, where transmitting full-resolution images is impractical and often unne

EgoEsportsQA: An Egocentric Video Benchmark for Perception and Reasoning in Esports

Model ReleasesDGX agent

arXiv:2604.12320v1 Announce Type: cross Abstract: While video large language models (Video-LLMs) excel in understanding slow-paced, real-world egocentric videos, their capabilities in high-velocity, i

El Agente Quntur: A research collaborator agent for quantum chemistry

AgentsDGX agent

arXiv:2602.04850v2 Announce Type: replace-cross Abstract: Quantum chemistry is a foundational enabling tool for the fields of chemistry, materials science, computational biology and others. Despite of

Elastic Net Regularization and Gabor Dictionary for Classification of Heart Sound Signals using Deep Learning

ResearchDGX agent

arXiv:2604.12483v1 Announce Type: cross Abstract: In this article, we propose the optimization of the resolution of time-frequency atoms and the regularization of fitting models to obtain better repre

EMBER: Autonomous Cognitive Behaviour from Learned Spiking Neural Network Dynamics in a Hybrid LLM Architecture

AgentsDGX agent

arXiv:2604.12167v1 Announce Type: new Abstract: We present (Experience-Modulated Biologically-inspired Emergent Reasoning), a hybrid cognitive architecture that reorganises the relationship between la

Enabling Ultra-Fast Cardiovascular Imaging Across Heterogeneous Clinical Environments with A Generalist Foundation Model and Multimodal Database

ResearchDGX agent

arXiv:2512.21652v2 Announce Type: replace-cross Abstract: Multimodal cardiovascular magnetic resonance (CMR) imaging provides comprehensive and non-invasive insights into cardiovascular disease (CVD)

Enhancing Clustering: An Explainable Approach via Filtered Patterns

ApplicationsDGX agent

arXiv:2604.12460v1 Announce Type: new Abstract: Machine learning has become a central research area, with increasing attention devoted to explainable clustering, also known as conceptual clustering, w

Enhancing Text-to-Image Diffusion Transformer via Split-Text Conditioning

ResearchDGX agent

arXiv:2505.19261v2 Announce Type: replace-cross Abstract: Current text-to-image diffusion generation typically employs complete-text conditioning. Due to the intricate syntax, diffusion transformers (

Euler-inspired Decoupling Neural Operator for Efficient Pansharpening

SafetyDGX agent

arXiv:2604.12463v1 Announce Type: cross Abstract: Pansharpening aims to synthesize high-resolution multispectral (HR-MS) images by fusing the spatial textures of panchromatic (PAN) images with the spe

Evaluating Language Models for Harmful Manipulation

Local AiDGX agent

arXiv:2603.25326v4 Announce Type: replace Abstract: Interest in the concept of AI-driven harmful manipulation is growing, yet current approaches to evaluating it are limited. This paper introduces a f

Evaluating LLM-Generated ACSL Annotations for Formal Verification

Model ReleasesDGX agent

arXiv:2602.13851v3 Announce Type: replace-cross Abstract: Formal specifications are crucial for building verifiable and dependable software systems, yet generating accurate and verifiable specificatio

Evaluating Relational Reasoning in LLMs with REL

Model ReleasesDGX agent

arXiv:2604.12176v1 Announce Type: new Abstract: Relational reasoning is the ability to infer relations that jointly bind multiple entities, attributes, or variables. This ability is central to scienti

Evaluating the Limitations of Protein Sequence Representations for Parkinson's Disease Classification

SafetyDGX agent

arXiv:2604.11852v1 Announce Type: cross Abstract: The identification of reliable molecular biomarkers for Parkinson's disease remains challenging due to its multifactorial nature. Although protein seq

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

SafetyDGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

FaCT: Faithful Concept Traces for Explaining Neural Network Decisions

Model ReleasesDGX agent

arXiv:2510.25512v2 Announce Type: replace-cross Abstract: Deep networks have shown remarkable performance across a wide range of tasks, yet getting a global concept-level understanding of how they fun

Fast AI Model Partition for Split Learning over Edge Networks

HardwareDGX agent

arXiv:2507.01041v4 Announce Type: replace-cross Abstract: Split learning (SL) is a distributed learning paradigm that can enable computation-intensive artificial intelligence (AI) applications by part

FAST-DIPS: Adjoint-Free Analytic Steps and Hard-Constrained Likelihood Correction for Diffusion-Prior Inverse Problems

Model ReleasesDGX agent

arXiv:2603.01591v2 Announce Type: replace-cross Abstract: Training-free diffusion priors enable inverse-problem solvers without retraining, but for nonlinear forward operators data consistency often r

FastGrasp: Learning-based Whole-body Control method for Fast Dexterous Grasping with Mobile Manipulators

ApplicationsDGX agent

arXiv:2604.12879v1 Announce Type: cross Abstract: Fast grasping is critical for mobile robots in logistics, manufacturing, and service applications. Existing methods face fundamental challenges in imp

Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces

Model ReleasesDGX agent

arXiv:2604.11996v1 Announce Type: cross Abstract: Should we trust Large Language Models (LLMs) with high accuracy? LLMs achieve high accuracy on reasoning benchmarks, but correctness alone does not re

Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data

HardwareDGX agent

arXiv:2503.10676v2 Announce Type: replace-cross Abstract: We study the efficacy of fine-tuning Large Language Models (LLMs) for the specific task of report (government archives, news, intelligence rep

FlowPlan-G2P: A Structured Generation Framework for Transforming Scientific Papers into Patent Descriptions

ApplicationsDGX agent

arXiv:2601.02589v2 Announce Type: replace-cross Abstract: Over 3.5 million patents are filed annually, with drafting patent descriptions requiring deep technical and legal expertise. Transforming scie

Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models

Model ReleasesDGX agent

arXiv:2506.14092v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in decision-support systems for high-stakes domains such as hiring and university admissions,

From edges to meaning: Semantic line sketches as a cognitive scaffold for ancient pictograph invention

ResearchDGX agent

arXiv:2604.12865v1 Announce Type: new Abstract: Humans readily recognize objects from sparse line drawings, a capacity that appears early in development and persists across cultures, suggesting neural

From Kinematics to Dynamics: Learning to Refine Hybrid Plans for Physically Feasible Execution

ResearchDGX agent

arXiv:2604.12474v1 Announce Type: cross Abstract: In many robotic tasks, agents must traverse a sequence of spatial regions to complete a mission. Such problems are inherently mixed discrete-continuou

From Plan to Action: How Well Do Agents Follow the Plan?

Model ReleasesDGX agent

arXiv:2604.12147v1 Announce Type: cross Abstract: Agents aspire to eliminate the need for task-specific prompt crafting through autonomous reason-act-observe loops. Still, they are commonly instructed

Frontier-Eng: Benchmarking Self-Evolving Agents on Real-World Engineering Tasks with Generative Optimization

Model ReleasesDGX agent

arXiv:2604.12290v1 Announce Type: new Abstract: Current LLM agent benchmarks, which predominantly focus on binary pass/fail tasks such as code generation or search-based question answering, often negl

FRTSearch: Unified Detection and Parameter Inference of Fast Radio Transients using Instance Segmentation

Model ReleasesDGX agent

arXiv:2604.12344v1 Announce Type: cross Abstract: The exponential growth of data from modern radio telescopes presents a significant challenge to traditional single-pulse search algorithms, which are

Fully Homomorphic Encryption on Llama 3 model for privacy preserving LLM inference

Model ReleasesDGX agent

arXiv:2604.12168v1 Announce Type: cross Abstract: The applications of Generative Artificial Intelligence (GenAI) and their intersections with data-driven fields, such as healthcare, finance, transport

GAM: Hierarchical Graph-based Agentic Memory for LLM Agents

AgentsDGX agent

arXiv:2604.12285v1 Announce Type: new Abstract: To sustain coherent long-term interactions, Large Language Model (LLM) agents must navigate the tension between acquiring new information and retaining

GCA Framework: A Gulf-Grounded Dataset and Agentic Pipeline for Climate Decision Support

Model ReleasesDGX agent

arXiv:2604.12306v1 Announce Type: cross Abstract: Climate decision-making in the Gulf increasingly demands systems that can translate heterogeneous scientific and policy evidence into actionable guida

GeM-EA: A Generative and Meta-learning Enhanced Evolutionary Algorithm for Streaming Data-Driven Optimization

Model ReleasesDGX agent

arXiv:2604.12336v1 Announce Type: cross Abstract: Streaming Data-Driven Optimization (SDDO) problems arise in many applications where data arrive continuously and the optimization environment evolves

Generative Modeling Enables Molecular Structure Retrieval from Coulomb Explosion Imaging

ResearchDGX agent

arXiv:2511.00179v2 Announce Type: replace-cross Abstract: Capturing the structural changes that molecules undergo during chemical reactions in real space and time is a long-standing dream and an essen

GeoPl@ntNet: A Platform for Exploring Essential Biodiversity Variables

ResearchDGX agent

arXiv:2511.13790v2 Announce Type: replace-cross Abstract: This paper describes GeoPl@ntNet, an interactive web application designed to make Essential Biodiversity Variables accessible and understandab

GF-Score: Certified Class-Conditional Robustness Evaluation with Fairness Guarantees

Model ReleasesDGX agent

arXiv:2604.12757v1 Announce Type: cross Abstract: Adversarial robustness is essential for deploying neural networks in safety-critical applications, yet standard evaluation methods either require expe

Global optimization tailored for graphics processing units: Complete and rigorous search for large-scale nonlinear minimization

Model ReleasesDGX agent

arXiv:2507.01770v4 Announce Type: replace-cross Abstract: This paper introduces a numerical method to enclose the global minimum of a nonlinear function subject to simple bounds on the variables. Usin

GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses

Model ReleasesDGX agent

arXiv:2604.11924v1 Announce Type: new Abstract: While LLMs hold significant potential to transform scientific research, we advocate for their use to augment and empower researchers rather than to auto

GRACE: A Dynamic Coreset Selection Framework for Large Language Model Optimization

ResearchDGX agent

arXiv:2604.11810v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in natural language understanding and generation. However, their immense number

GTCN-G: A Residual Graph-Temporal Fusion Network for Imbalanced Intrusion Detection

Model ReleasesDGX agent

arXiv:2510.07285v3 Announce Type: replace-cross Abstract: The escalating complexity of network threats and the inherent class imbalance in traffic data present formidable challenges for modern Intrusi

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

SafetyDGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

Heuristic Classification of Thoughts Prompting (HCoT): Integrating Expert System Heuristics for Structured Reasoning into Large Language Models

TutorialsDGX agent

arXiv:2604.12390v1 Announce Type: new Abstract: This paper addresses two limitations of large language models (LLMs) in solving complex problems: (1) their reasoning processes exhibit Bayesian-like st

HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST

SafetyDGX agent

arXiv:2509.19742v4 Announce Type: replace-cross Abstract: Zero-shot Dialog State Tracking (zs-DST) is essential for enabling Task-Oriented Dialog Systems (TODs) to generalize to new domains without co

HiFiNet: Hierarchical Fault Identification in Wireless Sensor Networks via Edge-Based Classification and Graph Aggregation

ResearchDGX agent

arXiv:2511.17537v2 Announce Type: replace-cross Abstract: Wireless Sensor Networks (WSN) are the backbone of essential monitoring applications, but their deployment in unfavourable conditions increase

HintMR: Eliciting Stronger Mathematical Reasoning in Small Language Models

Local AiDGX agent

arXiv:2604.12229v1 Announce Type: new Abstract: Small language models (SLMs) often struggle with complex mathematical reasoning due to limited capacity to maintain long chains of intermediate steps an

How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm

Model ReleasesDGX agent

arXiv:2604.12250v1 Announce Type: new Abstract: This study examines how model-specific characteristics of Large Language Model (LLM) agents, including internal alignment, shape the effect of memory on

How Transformers Learn to Plan via Multi-Token Prediction

SafetyDGX agent

arXiv:2604.11912v1 Announce Type: cross Abstract: While next-token prediction (NTP) has been the standard objective for training language models, it often struggles to capture global structure in reas

← Previous
1…327328329330331…354
Next →