AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
28 Apr 2026

Interpretable Physics-Informed Load Forecasting for U.S. Grid Resilience: SHAP-Guided Ensemble Validation in Hybrid Deep Learning Under Extreme Weather

ResearchDGX agent

arXiv:2604.23500v1 Announce Type: cross Abstract: Accurate short-term electricity load forecasting is a cornerstone of U.S. grid reliability; however, prevailing deep learning models remain opaque, li

Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics

SafetyDGX agent

arXiv:2604.22832v1 Announce Type: cross Abstract: Microscopy-based phenotypic profiling is scalable for drug discovery but lacks the mechanistic depth of transcriptomics, which remains costly and scar

IntrAgent: An LLM Agent for Content-Grounded Information Retrieval through Literature Review


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.22861v1 Announce Type: cross Abstract: Scientific research relies on accurate information retrieval from literature to support analytical decisions. In this work, we introduce a new task, I

Inverting Foundation Models of Brain Function with Simulation-Based Inference

TutorialsDGX agent

arXiv:2604.23865v1 Announce Type: cross Abstract: Foundation models of brain activity promise a new frontier for in silico neuroscience by emulating neural responses to complex stimuli across tasks an

Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens

SafetyDGX agent

arXiv:2508.01191v5 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has been shown to be effective in eliciting structured reasoning (i.e., CoT reasoning) from large language models (

Isotonic Layer: A Unified Framework for Recommendation Calibration and Debiasing

SafetyDGX agent

arXiv:2603.06589v2 Announce Type: replace-cross Abstract: Model calibration and debiasing are fundamental yet operationally expensive challenges in large-scale recommendation systems. Existing approac

Jailbreaking Frontier Foundation Models Through Intention Deception

Model ReleasesDGX agent

arXiv:2604.24082v1 Announce Type: cross Abstract: Large (vision-)language models exhibit remarkable capability but remain highly susceptible to jailbreaking. Existing safety training approaches aim to

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

Model ReleasesDGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology

Model ReleasesDGX agent

arXiv:2604.24645v1 Announce Type: cross Abstract: The development of practical (multimodal) large language model assistants for Korean weather forecasters is hindered by the absence of a multidimensio

K-Score: Kalman Filter as a Principled Alternative to Reward Normalization in Reinforcement Learning

SafetyDGX agent

arXiv:2604.23056v1 Announce Type: cross Abstract: We propose a simple yet effective alternative to reward normalization in policy gradient reinforcement learning by integrating a 1D Kalman filter for

K-SENSE: A Knowledge-Guided Self-Augmented Encoder for Neuro-Semantic Evaluation of Mental Health Conditions on Social Media

ResearchDGX agent

arXiv:2604.23493v1 Announce Type: cross Abstract: Early detection of mental health conditions, particularly stress and depression, from social media text remains a challenging open problem in computat

KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning

AgentsDGX agent

arXiv:2604.22779v1 Announce Type: cross Abstract: Enabling large language models (LLMs) to appropriately abstain from answering questions beyond their knowledge is crucial for mitigating hallucination

KLong: Training LLM Agent for Extremely Long-horizon Tasks

Model ReleasesDGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

Knee-xRAI: An Explainable AI Framework for Automatic Kellgren-Lawrence Grading of Knee Osteoarthritis

ResearchDGX agent

arXiv:2604.23435v1 Announce Type: cross Abstract: Radiographic grading of knee osteoarthritis (KOA) with the Kellgren-Lawrence (KL) system is limited by inter-reader variability and the opacity of cur

Knowledge Lever Risk Management for Software Engineering: A Stochastic Framework for Mitigating Knowledge Loss

SafetyDGX agent

arXiv:2604.23257v1 Announce Type: cross Abstract: Software engineering (SE) organizations operate in a knowledge-intensive domain where critical assets -- architectural expertise, design rationale, an

KOMBO: Korean Character Representations Based on the Combination Rules of Subcharacters

ResearchDGX agent

arXiv:2604.23948v1 Announce Type: cross Abstract: The Korean writing system, extit{Hangeul}, has a unique character representation rigidly following the invention principles recorded in extit{Hunminje

Kwai Summary Attention Technical Report

Local AiDGX agent

arXiv:2604.24432v1 Announce Type: cross Abstract: Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic unde

Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting

ResearchDGX agent

arXiv:2506.19089v5 Announce Type: replace-cross Abstract: We introduce StorySim, a programmable framework for synthetically generating stories to evaluate the theory of mind (ToM) and world modeling (

Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

Model ReleasesDGX agent

arXiv:2509.06337v2 Announce Type: replace Abstract: Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-

Latency and Cost of Multi-Agent Intelligent Tutoring at Scale

Model ReleasesDGX agent

arXiv:2604.24110v1 Announce Type: cross Abstract: Multi-agent LLM tutoring systems improve response quality through agent specialization, but each student query triggers several concurrent API calls w

Latent-Hysteresis Graph ODEs: Modeling Coupled Topology-Feature Evolution via Continuous Phase Transitions

ApplicationsDGX agent

arXiv:2604.24293v1 Announce Type: cross Abstract: Graph neural ordinary differential equations (Graph ODEs) extend graph learning from discrete message-passing layers to continuous-time representation

Layer Embedding Deep Fusion Graph Neural Network

ResearchDGX agent

arXiv:2604.23324v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have demonstrated impressive performance in learning representations from graph-structured data. However, their message-p

Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models

Model ReleasesDGX agent

arXiv:2604.24542v1 Announce Type: cross Abstract: Large language models deployed at runtime can misbehave in ways that clean-data validation cannot anticipate: training-time backdoors lie dormant unti

LearnAlign: Data Selection for LLM Reinforcement Learning with Improved Gradient Alignment

Model ReleasesDGX agent

arXiv:2506.11480v4 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a key technique for enhancing LLMs' reasoning abilities, yet its data ineffic

Learn&Drop: Fast Learning of CNNs based on Layer Dropping

TutorialsDGX agent

arXiv:2604.23403v1 Announce Type: cross Abstract: This paper proposes a new method to improve the training efficiency of deep convolutional neural networks. During training, the method evaluates score

Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning

Model ReleasesDGX agent

arXiv:2604.22770v1 Announce Type: cross Abstract: Most digital language learning curricula rely on discrete-item quizzes that test recall rather than applied conversational proficiency. When progressi

Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain

Model ReleasesDGX agent

arXiv:2509.10546v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in finance, where unsafe behavior can lead to serious regulatory risks. However, most r

Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs

SafetyDGX agent

arXiv:2509.00084v2 Announce Type: replace-cross Abstract: Test-time scaling (TTS) has gained widespread attention for enhancing LLM reasoning. Existing approaches such as Best-of-N and majority voting

Learning to Rotate: Temporal and Semantic Rotary Encoding for Sequential Modeling

ApplicationsDGX agent

arXiv:2604.24717v1 Announce Type: new Abstract: Every Transformer architecture dedicates enormous capacity to learning rich representations in semantic embedding space -- yet the rotation manifold act

Learning to Route Queries to Heads for Attention-based Re-ranking with Large Language Models

TutorialsDGX agent

arXiv:2604.24608v1 Announce Type: cross Abstract: Large Language Models (LLMs) have recently been explored as fine-grained zero-shot re-rankers by leveraging attention signals to estimate document rel

Learning to Think from Multiple Thinkers

TutorialsDGX agent

arXiv:2604.24737v1 Announce Type: cross Abstract: We study learning with Chain-of-Thought (CoT) supervision from multiple thinkers, all of whom provide correct but possibly systematically different so

LEGO: An LLM Skill-Based Front-End Design Generation Platform

Model ReleasesDGX agent

arXiv:2604.23355v1 Announce Type: new Abstract: Existing LLM-based EDA agents are often isolated task-specific systems. This leads to repeated engineering effort and limited reuse of successful design

Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Application

Model ReleasesDGX agent

arXiv:2604.24636v1 Announce Type: cross Abstract: On-device Small Language Models (SLMs) promise fully offline, private AI experiences for mobile users (no cloud dependency, no data leaving the device

Leveraging Human Feedback for Semantically-Relevant Skill Discovery

ResearchDGX agent

arXiv:2604.24127v1 Announce Type: cross Abstract: Unsupervised skill discovery in reinforcement learning aims to intrinsically motivate agents to discover diverse and useful behaviours. However, uncon

Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study

Model ReleasesDGX agent

arXiv:2604.24678v1 Announce Type: cross Abstract: Large language models (LLMs) perform strongly on general-purpose code generation, yet their applicability to enterprise domain-specific languages (DSL

Lightweight and Production-Ready PDF Visual Element Parsing

Model ReleasesDGX agent

arXiv:2604.23276v1 Announce Type: cross Abstract: PDF documents contain critical visual elements such as figures, tables, and forms whose accurate extraction is essential for document understanding an

LLM-Assisted Op-Amp Behavioral-Level Design via Agentic Human-Mimicking Reasoning

Model ReleasesDGX agent

arXiv:2601.21321v2 Announce Type: replace Abstract: This paper proposes White-Op, an operational amplifier (op-amp) behavioral-level parameter design framework assisted by the human-mimicking reasonin

LLM-Auction: Generative Auction towards LLM-Native Advertising

SafetyDGX agent

arXiv:2512.10551v2 Announce Type: replace-cross Abstract: The commercialization of LLM applications is the next frontier in online advertising, with LLM-native advertising emerging as a promising para

LLM-Augmented Traffic Signal Control with LSTM-Based Traffic State Prediction and Safety-Constrained Decision Support

SafetyDGX agent

arXiv:2604.23902v1 Announce Type: new Abstract: Traffic signal control is a critical task in intelligent transportation systems, yet conventional fixed-time and rule-based methods often struggle to ad

LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

Model ReleasesDGX agent

arXiv:2604.23970v1 Announce Type: new Abstract: Indoor navigation remains a critical accessibility challenge for the blind and low-vision (BLV) individuals, as existing solutions rely on costly per-bu

LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews

Model ReleasesDGX agent

arXiv:2511.12635v2 Announce Type: replace-cross Abstract: Context: Large language models (LLMs) are increasingly used to screen literature for systematic reviews (SRs), but the standard confusion-matr

LLMs Reading the Rhythms of Daily Life: Aligned Understanding for Behavior Prediction and Generation

SafetyDGX agent

arXiv:2604.23578v1 Announce Type: cross Abstract: Human daily behavior unfolds as complex sequences shaped by intentions, preferences, and context. Effectively modeling these behaviors is crucial for

LoFi: Location-Aware Fine-Grained Representation Learning for Chest X-ray

ResearchDGX agent

arXiv:2603.19451v2 Announce Type: replace-cross Abstract: Fine-grained representation learning is crucial for retrieval and phrase grounding in chest X-rays, where clinically relevant findings are oft

Lost in Decoding? Reproducing and Stress-Testing the Look-Ahead Prior in Generative Retrieval

Model ReleasesDGX agent

arXiv:2604.23396v1 Announce Type: cross Abstract: Generative retrieval (GR) ranks documents by autoregressively generating document identifiers. Because many GR methods rely on trie-constrained beam s

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation

ResearchDGX agent

arXiv:2603.17717v3 Announce Type: replace-cross Abstract: Supervised detection of network attacks has always been a critical part of network intrusion detection systems (NIDS). Nowadays, in a pivotal

MAE-Based Self-Supervised Pretraining for Data-Efficient Medical Image Segmentation Using nnFormer

TutorialsDGX agent

arXiv:2604.22854v1 Announce Type: cross Abstract: Transformer architectures, including nnFormer,have demonstrated promising results in volumetric medical image segmentation by being able to capture lo

Mapping License Plate Recoverability Under Extreme Viewing Angles for Oppor-tunistic Urban Sensing

Model ReleasesDGX agent

arXiv:2604.23814v1 Announce Type: cross Abstract: Urban environments contain many imaging sensors built for specific purposes, including ATM, body-worn, CCTV, and dashboard cameras. Under the opportun

MarketBench: Evaluating AI Agents as Market Participants

Model ReleasesDGX agent

arXiv:2604.23897v1 Announce Type: new Abstract: Markets are a promising way to coordinate AI agent activity for similar reasons to those used to justify markets more broadly. In order to effectively p

MEASER: Malware embedding attacks on open-source LLMs

Model ReleasesDGX agent

arXiv:2510.10486v2 Announce Type: replace-cross Abstract: Open-source large language models (LLMs) have demonstrated considerable dominance over proprietary LLMs in resolving neural processing tasks,

Measuring Successful Cooperation in Human-AI Teamwork: Development and Validation of the Perceived Cooperativity and Teaming Perception Scales

AgentsDGX agent

arXiv:2604.24461v1 Announce Type: cross Abstract: As human-AI cooperation becomes increasingly prevalent, reliable instruments for assessing the subjective quality of cooperative human-AI interaction

Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings

Model ReleasesDGX agent

arXiv:2604.23130v1 Announce Type: cross Abstract: Large language models (LLMs) can still be jailbroken into producing harmful outputs despite safety alignment. Existing attacks show this vulnerability

MedSpeak: A Knowledge Graph-Aided ASR Error Correction Framework for Spoken Medical QA

ResearchDGX agent

arXiv:2602.00981v2 Announce Type: replace-cross Abstract: Spoken question-answering (SQA) systems relying on automatic speech recognition (ASR) often struggle with accurately recognizing medical termi

MegaScale-Data: Scaling Dataloader for Multisource Large Foundation Model Training

ResearchDGX agent

arXiv:2504.09844v4 Announce Type: replace-cross Abstract: Modern frameworks for training large foundation models (LFMs) employ dataloaders in a data-parallel manner, with each loader processing a disj

MEMCoder: Multi-dimensional Evolving Memory for Private-Library-Oriented Code Generation

Model ReleasesDGX agent

arXiv:2604.24222v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at general code generation, but their performance drops sharply in enterprise settings that rely on internal privat

MemeScouts@LT-EDI 2026: Asking the Right Questions -- Prompted Weak Supervision for Meme Hate Speech Detection

ResearchDGX agent

arXiv:2604.24179v1 Announce Type: cross Abstract: Detecting hate speech in memes is challenging due to their multimodal nature and subtle, culturally grounded cues such as sarcasm and context. While r

MERIT: Modular Framework for Multimodal Misinformation Detection with Web-Grounded Reasoning

SafetyDGX agent

arXiv:2510.17590v2 Announce Type: replace Abstract: We present MERIT, an inference-time modular framework for multimodal misinformation detection that decomposes verification into four specialized mod

MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation

Model ReleasesDGX agent

arXiv:2511.14967v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown great promise in generating structured diagrams from natural language descriptions, particularly Merma

Meta-Aligner: Bidirectional Preference-Policy Optimization for Multi-Objective LLMs Alignment

SafetyDGX agent

arXiv:2604.24178v1 Announce Type: cross Abstract: Multi-Objective Alignment aims to align Large Language Models (LLMs) with diverse and often conflicting human values by optimizing multiple objectives

Meta-CoT: Enhancing Granularity and Generalization in Image Editing

Model ReleasesDGX agent

arXiv:2604.24625v1 Announce Type: cross Abstract: Unified multi-modal understanding/generative models have shown improved image editing performance by incorporating fine-grained understanding into the

Meta-Ensemble Learning with Diverse Data Splits for Improved Respiratory Sound Classification

Model ReleasesDGX agent

arXiv:2604.24096v1 Announce Type: cross Abstract: Training reliable respiratory sound classification models remains challenging due to the limited size and subject diversity of datasets. Ensemble meth

← Previous
1…299300301302303…354
Next →