AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
15 Apr 2026

HSG-12M: A Large-Scale Benchmark of Spatial Multigraphs from the Energy Spectra of Non-Hermitian Crystals

Model ReleasesDGX agent

arXiv:2506.08618v4 Announce Type: replace-cross Abstract: AI is transforming scientific research by revealing new ways to understand complex physical systems, but its impact remains constrained by the

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

SafetyDGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

Human-Inspired Context-Selective Multimodal Memory for Social Robots

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12081v1 Announce Type: new Abstract: Memory is fundamental to social interaction, enabling humans to recall meaningful past experiences and adapt their behavior accordingly based on the con

Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance

SafetyDGX agent

arXiv:2511.21356v2 Announce Type: replace-cross Abstract: Adversarial Inverse Reinforcement Learning (AIRL) has shown promise in addressing the sparse reward problem in reinforcement learning (RL) by

IAD-Unify: A Region-Grounded Unified Model for Industrial Anomaly Segmentation, Understanding, and Generation

Model ReleasesDGX agent

arXiv:2604.12440v1 Announce Type: cross Abstract: Real-world industrial inspection requires not only localizing defects, but also explaining them in natural language and generating controlled defect e

IDEA: An Interpretable and Editable Decision-Making Framework for LLMs via Verbal-to-Numeric Calibration

Model ReleasesDGX agent

arXiv:2604.12573v1 Announce Type: new Abstract: Large Language Models are increasingly deployed for decision-making, yet their adoption in high-stakes domains remains limited by miscalibrated probabil

Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space

Model ReleasesDGX agent

arXiv:2604.12016v1 Announce Type: new Abstract: Large language models map semantically related prompts to similar internal representations -- a phenomenon interpretable as attractor-like dynamics. We

Improved particle swarm optimization algorithm: multi-target trajectory optimization for swarm drones

Model ReleasesDGX agent

arXiv:2507.13647v2 Announce Type: replace-cross Abstract: Real-time trajectory planning for unmanned aerial vehicles (UAVs) in dynamic environments remains a key challenge due to high computational de

IMSE: Intrinsic Mixture of Spectral Experts Fine-tuning for Test-Time Adaptation

Model ReleasesDGX agent

arXiv:2603.07926v3 Announce Type: replace-cross Abstract: Test-time adaptation (TTA) has been widely explored to prevent performance degradation when test data differ from the training distribution. H

INDOTABVQA: A Benchmark for Cross-Lingual Table Understanding in Bahasa Indonesia Documents

Model ReleasesDGX agent

arXiv:2604.11970v1 Announce Type: cross Abstract: We introduce INDOTABVQA, a benchmark for evaluating cross-lingual Table Visual Question Answering (VQA) on real-world document images in Bahasa Indone

INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT

Model ReleasesDGX agent

arXiv:2512.14732v2 Announce Type: replace-cross Abstract: Incidental findings in CT scans, though often benign, can have significant clinical implications and should be reported following established

Information-Theoretic Optimization for Task-Adapted Compressed Sensing Magnetic Resonance Imaging

ResearchDGX agent

arXiv:2604.12709v1 Announce Type: cross Abstract: Task-adapted compressed sensing magnetic resonance imaging (CS-MRI) is emerging to address the specific demands of downstream clinical tasks with sign

Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following

AgentsDGX agent

arXiv:2510.14420v4 Announce Type: replace-cross Abstract: Language models often struggle to follow multi-constraint instructions that are crucial for real-world applications. Existing reinforcement le

Intelligent ROI-Based Vehicle Counting Framework for Automated Traffic Monitoring

Model ReleasesDGX agent

arXiv:2604.12470v1 Announce Type: new Abstract: Accurate vehicle counting through video surveillance is crucial for efficient traffic management. However, achieving high counting accuracy while ensuri

Interpretable DNA Sequence Classification via Dynamic Feature Generation in Decision Trees

ResearchDGX agent

arXiv:2604.12060v1 Announce Type: cross Abstract: The analysis of DNA sequences has become critical in numerous fields, from evolutionary biology to understanding gene regulation and disease mechanism

Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety

Model ReleasesDGX agent

arXiv:2604.12311v1 Announce Type: cross Abstract: The emergence of vibe coding, a paradigm where non-technical users instruct Large Language Models (LLMs) to generate executable codes via natural lang

JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence

ResearchDGX agent

arXiv:2510.23538v2 Announce Type: replace Abstract: The scope of neural code intelligence is rapidly expanding beyond text-based source code to encompass the rich visual outputs that programs generate

Joint Flashback Adaptation for Forgetting-Resistant Instruction Tuning

TutorialsDGX agent

arXiv:2505.15467v2 Announce Type: replace-cross Abstract: Large language models have achieved remarkable success in various tasks. However, it is challenging for them to learn new tasks incrementally

KG-Hopper: Empowering Compact Open LLMs with Knowledge Graph Reasoning via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.21440v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate impressive natural language capabilities but often struggle with knowledge-intensive reasoning tasks.

KG-Reasoner: A Reinforced Model for End-to-End Multi-Hop Knowledge Graph Reasoning

ResearchDGX agent

arXiv:2604.12487v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit strong abilities in natural language understanding and generation, yet they struggle with knowledge-intensive rea

KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidance

Model ReleasesDGX agent

arXiv:2604.12627v1 Announce Type: new Abstract: RLVR improves reasoning in large language models, but its effectiveness is often limited by severe reward sparsity on hard problems. Recent hint-based R

KumoRFM-2: Scaling Foundation Models for Relational Learning

Model ReleasesDGX agent

arXiv:2604.12596v1 Announce Type: cross Abstract: We introduce KumoRFM-2, the next iteration of a pre-trained foundation model for relational data. KumoRFM-2 supports in-context learning as well as fi

Large Language Models are Powerful Electronic Health Record Encoders

Model ReleasesDGX agent

arXiv:2502.17403v5 Announce Type: replace-cross Abstract: Electronic Health Records (EHRs) offer considerable potential for clinical prediction, but their complexity and heterogeneity challenge tradit

LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety

Model ReleasesDGX agent

arXiv:2604.12710v1 Announce Type: cross Abstract: Large language models (LLMs) often demonstrate strong safety performance in high-resource languages, yet exhibit severe vulnerabilities when queried i

Latent patterns of urban mixing in mobility analysis across five global cities

ResearchDGX agent

arXiv:2604.12202v1 Announce Type: new Abstract: This study leverages large-scale travel surveys for over 200,000 residents across Boston, Chicago, Hong Kong, London, and Sao Paulo. With rich individua

Latent Planning Emerges with Scale

Model ReleasesDGX agent

arXiv:2604.12493v1 Announce Type: cross Abstract: LLMs can perform seemingly planning-intensive tasks, like writing coherent stories or functioning code, without explicitly verbalizing a plan; however

LatentRefusal: Latent-Signal Refusal for Unanswerable Text-to-SQL Queries

SafetyDGX agent

arXiv:2601.10398v3 Announce Type: replace Abstract: In LLM-based text-to-SQL systems, unanswerable and underspecified user queries may generate not only incorrect text but also executable programs tha

League of LLMs: A Benchmark-Free Paradigm for Mutual Evaluation of Large Language Models

Model ReleasesDGX agent

arXiv:2507.22359v4 Announce Type: replace Abstract: Although large language models (LLMs) have shown exceptional capabilities across a wide range of tasks, reliable evaluation remains a critical chall

Learning Chain Of Thoughts Prompts for Predicting Entities, Relations, and even Literals on Knowledge Graphs

ResearchDGX agent

arXiv:2604.12651v1 Announce Type: cross Abstract: Knowledge graph embedding (KGE) models perform well on link prediction but struggle with unseen entities, relations, and especially literals, limiting

Learning the Value of Value Learning

AgentsDGX agent

arXiv:2511.17714v5 Announce Type: replace Abstract: Standard decision frameworks address uncertainty about facts but assume fixed options and values. We extend the Jeffrey-Bolker framework to model re

Leveraging Weighted Syntactic and Semantic Context Assessment Summary (wSSAS) Towards Text Categorization Using LLMs

Model ReleasesDGX agent

arXiv:2604.12049v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) for reliable, enterprise-grade analytics such as text categorization is often hindered by the stochastic natur

LIFE -- an energy efficient advanced continual learning agentic AI framework for frontier systems

AgentsDGX agent

arXiv:2604.12874v1 Announce Type: new Abstract: The rapid advancement of AI has changed the character of HPC usage such as dimensioning, provisioning, and execution. Not only has energy demand been am

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

SafetyDGX agent

arXiv:2604.13010v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, standard OPD requires a live teach

Lit2Vec: A Reproducible Workflow for Building a Legally Screened Chemistry Corpus from S2ORC for Downstream Retrieval and Text Mining

Model ReleasesDGX agent

arXiv:2604.12498v1 Announce Type: cross Abstract: We present Lit2Vec, a reproducible workflow for constructing and validating a chemistry corpus from the Semantic Scholar Open Research Corpus using co

LLM as Attention-Informed NTM and Topic Modeling as long-input Generation: Interpretability and long-Context Capability

ResearchDGX agent

arXiv:2510.03174v2 Announce Type: replace-cross Abstract: Topic modeling aims to produce interpretable topic representations and topic--document correspondences from corpora, but classical neural topi

LLM-Based Automated Diagnosis Of Integration Test Failures At Google

ApplicationsDGX agent

arXiv:2604.12108v1 Announce Type: cross Abstract: Integration testing is critical for the quality and reliability of complex software systems. However, diagnosing their failures presents significant c

LLM-Guided Prompt Evolution for Password Guessing

Model ReleasesDGX agent

arXiv:2604.12601v1 Announce Type: cross Abstract: Passwords still remain a dominant authentication method, yet their security is routinely subverted by predictable user choices and large-scale credent

LLM-Guided Semantic Bootstrapping for Interpretable Text Classification with Tsetlin Machines

TutorialsDGX agent

arXiv:2604.12223v1 Announce Type: cross Abstract: Pretrained language models (PLMs) like BERT provide strong semantic representations but are costly and opaque, while symbolic models such as the Tsetl

LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks

ApplicationsDGX agent

arXiv:2604.12096v1 Announce Type: new Abstract: On online advertising platforms, newly introduced promotional ads face the cold-start problem, as they lack sufficient user feedback for model training.

LLMs Struggle with Abstract Meaning Comprehension More Than Expected

ResearchDGX agent

arXiv:2604.12018v1 Announce Type: cross Abstract: Understanding abstract meanings is crucial for advanced language comprehension. Despite extensive research, abstract words remain challenging due to t

Local-Splitter: A Measurement Study of Seven Tactics for Reducing Cloud LLM Token Usage on Coding-Agent Workloads

Local AiDGX agent

arXiv:2604.12301v1 Announce Type: cross Abstract: We present a systematic measurement study of seven tactics for reducing cloud LLM token usage when a small local model can act as a triage layer in fr

LogicEval: A Systematic Framework for Evaluating Automated Repair Techniques for Logical Vulnerabilities in Real-World Software

SafetyDGX agent

arXiv:2604.12994v1 Announce Type: cross Abstract: Logical vulnerabilities in software stem from flaws in program logic rather than memory safety, which can lead to critical security failures. Although

Long-Horizon Plan Execution in Large Tool Spaces through Entropy-Guided Branching

Model ReleasesDGX agent

arXiv:2604.12126v1 Announce Type: new Abstract: Large Language Models (LLMs) have significantly advanced tool-augmented agents, enabling autonomous reasoning via API interactions. However, executing m

Loop Corrections to the Training and Generalization Errors of Random Feature Models

ResearchDGX agent

arXiv:2604.12827v1 Announce Type: cross Abstract: We investigate random feature models in which neural networks sampled from a prescribed initialization ensemble are frozen and used as random features

Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain

AgentsDGX agent

arXiv:2510.05159v4 Announce Type: replace-cross Abstract: While finetuning AI agents on interaction data -- such as web browsing or tool use -- improves their capabilities, it also introduces critical

MAML-KT: Addressing Cold Start Problem in Knowledge Tracing for New Students via Few-Shot Model-Agnostic Meta Learning

ResearchDGX agent

arXiv:2603.00137v2 Announce Type: replace-cross Abstract: Knowledge tracing (KT) models are commonly evaluated by training on early interactions from all students and testing on later responses. While

Man and machine: artificial intelligence and judicial decision making

SafetyDGX agent

arXiv:2603.19042v4 Announce Type: replace Abstract: The integration of artificial intelligence (AI) technologies into judicial decision-making, particularly in pretrial, sentencing, and parole context

Mantis: A Foundation Model for Mechanistic Disease Forecasting

ApplicationsDGX agent

arXiv:2508.12260v5 Announce Type: replace Abstract: Infectious disease forecasting in novel outbreaks or low-resource settings is hampered by the need for large disease and covariate data sets, bespok

MAST: Mask-Guided Attention Mass Allocation for Training-Free Multi-Style Transfer

ResearchDGX agent

arXiv:2604.12281v1 Announce Type: cross Abstract: Style transfer aims to render a content image with the visual characteristics of a reference style while preserving its underlying semantic layout and

Mathematics Teachers Interactions with a Multi-Agent System for Personalized Problem Generation

AgentsDGX agent

arXiv:2604.12066v1 Announce Type: new Abstract: Large language models can increasingly adapt educational tasks to learners characteristics. In the present study, we examine a multi-agent teacher-in-th

Memory as Metabolism: A Design for Companion Knowledge Systems

Model ReleasesDGX agent

arXiv:2604.12034v1 Announce Type: new Abstract: Retrieval-Augmented Generation remains the dominant pattern for giving LLMs persistent memory, but a visible cluster of personal wiki-style memory archi

Mining Large Language Models for Low-Resource Language Data: Comparing Elicitation Strategies for Hausa and Fongbe

Model ReleasesDGX agent

arXiv:2604.12477v1 Announce Type: cross Abstract: Large language models (LLMs) are trained on data contributed by low-resource language communities, yet the linguistic knowledge encoded in these model

MISID: A Multimodal Multi-turn Dataset for Complex Intent Recognition in Strategic Deception Games

Model ReleasesDGX agent

arXiv:2604.12700v1 Announce Type: new Abstract: Understanding human intent in complex multi-turn interactions remains a fundamental challenge in human-computer interaction and behavioral analysis. Whi

Mixed-Density Diffuser: Efficient Planning with Non-Uniform Temporal Resolution

Model ReleasesDGX agent

arXiv:2510.23026v5 Announce Type: replace Abstract: Recent studies demonstrate that diffusion planners benefit from sparse-step planning over single-step planning. Training models to skip steps in the

Mobile GUI Agents under Real-world Threats: Are We There Yet?

Model ReleasesDGX agent

arXiv:2507.04227v2 Announce Type: replace-cross Abstract: Recent years have witnessed a rapid development of mobile GUI agents powered by large language models (LLMs), which can autonomously execute d

Modality-Native Routing in Agent-to-Agent Networks: A Multimodal A2A Protocol Extension

Model ReleasesDGX agent

arXiv:2604.12213v1 Announce Type: new Abstract: Preserving multimodal signals across agent boundaries is necessary for accurate cross-modal reasoning, but it is not sufficient. We show that modality-n

Modeling Co-Pilots for Text-to-Model Translation

AgentsDGX agent

arXiv:2604.12955v1 Announce Type: new Abstract: There is growing interest in leveraging large language models (LLMs) for text-to-model translation and optimization tasks. This paper aims to advance th

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

SafetyDGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

MoDora: Tree-Based Semi-Structured Document Analysis System

Local AiDGX agent

arXiv:2602.23061v3 Announce Type: replace-cross Abstract: Semi-structured documents integrate diverse interleaved data elements (e.g., tables, charts, hierarchical paragraphs) arranged in various and

MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization

SafetyDGX agent

arXiv:2604.12237v1 Announce Type: cross Abstract: In drug discovery, molecular optimization aims to iteratively refine a lead compound to improve molecular properties while preserving structural simil

← Previous
1…328329330331332…354
Next →