AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
1 May 2026

Interval Orders, Biorders and Credibility-limited Belief Revision

AgentsDGX agent

arXiv:2604.27156v1 Announce Type: new Abstract: Rational belief revision is commonly viewed as being based on a preference order between possible worlds, with the resulting new belief set being those

Investigating More Explainable and Partition-Free Compositionality Estimation for LLMs: A Rule-Generation Perspective

ResearchDGX agent

arXiv:2604.27340v1 Announce Type: new Abstract: Compositional generalization tests are often used to estimate the compositionality of LLMs. However, such tests have the following limitations: (1) they

Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.27724v1 Announce Type: new Abstract: Medical retrieval-augmented generation (RAG) systems typically operate on text chunks extracted from biomedical literature, discarding the rich visual c

ITS-Mina: A Harris Hawks Optimization-Based All-MLP Framework with Iterative Refinement and External Attention for Multivariate Time Series Forecasting

Model ReleasesDGX agent

arXiv:2604.27981v1 Announce Type: cross Abstract: Multivariate time series forecasting plays a pivotal role in numerous real-world applications, including financial analysis, energy management, and tr

Junk DNA Hypothesis: Pruning Small Pre-Trained Weights Irreversibly and Monotonically Impairs 'Difficult' Downstream Tasks in LLMs

ResearchDGX agent

arXiv:2310.02277v4 Announce Type: replace-cross Abstract: We present Junk DNA Hypothesis by adopting a novel task-centric angle for the pre-trained weights of large language models (LLMs). It has been

K2MUSE: A human lower-limb multimodal walking dataset spanning task and acquisition variability for rehabilitation robotics

Model ReleasesDGX agent

arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

Model ReleasesDGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

Knowledge Affordances for Hybrid Human-AI Information Seeking

AgentsDGX agent

arXiv:2604.27539v1 Announce Type: cross Abstract: As information ecosystems grow more heterogeneous, both humans and artificial agents increasingly face a simple yet unresolved question: when seeking

Knowledge Graph Representations for LLM-Based Policy Compliance Reasoning

SafetyDGX agent

arXiv:2604.27713v1 Announce Type: new Abstract: The risks posed by AI features are increasing as they are rapidly integrated into software applications. In response, regulations and standards for safe

Language Models Refine Mechanical Linkage Designs Through Symbolic Reflection and Modular Optimisation

Model ReleasesDGX agent

arXiv:2604.27962v1 Announce Type: new Abstract: Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematicall

Latent Adversarial Detection: Adaptive Probing of LLM Activations for Multi-Turn Attack Detection

ApplicationsDGX agent

arXiv:2604.28129v1 Announce Type: cross Abstract: Multi-turn prompt injection follows a known attack path -- trust-building, pivoting, escalation but text-level defenses miss covert attacks where indi

Leading Across the Spectrum of Human-AI Relationships: A Conceptual Framework for Increasingly Heterogeneous Teams

ResearchDGX agent

arXiv:2604.27392v1 Announce Type: new Abstract: What shapes a consequential decision when human and artificial intelligence work on it together? The answer is becoming harder to see. A decision may lo

Learning from Disagreement: Clinician Overrides as Implicit Preference Signals for Clinical AI in Value-Based Care

SafetyDGX agent

arXiv:2604.28010v1 Announce Type: cross Abstract: We reframe clinician overrides of clinical AI recommendations as implicit preference data - the same signal structure exploited by reinforcement learn

Learning Generalizable Multimodal Representations for Software Vulnerability Detection

Model ReleasesDGX agent

arXiv:2604.25711v2 Announce Type: replace-cross Abstract: Source code and its accompanying comments are complementary yet naturally aligned modalities-code encodes structural logic while comments capt

Learning Rate Engineering: From Coarse Single Parameter to Layered Evolution

Model ReleasesDGX agent

arXiv:2604.27295v1 Announce Type: new Abstract: Learning rate scheduling has evolved from the single global fixed rate of early SGD to sophisticated layer-wise adaptive strategies. We systematize this

Learning Rate Transfer in Normalized Transformers

SafetyDGX agent

arXiv:2604.27077v1 Announce Type: cross Abstract: The Normalized Transformer, or nGPT (arXiv:2410.01131) achieves impressive training speedups and does not require weight decay or learning rate warmup

Learning to Aggregate Zero-Shot LLM Agents for Corporate Disclosure Classification

ResearchDGX agent

arXiv:2603.20965v2 Announce Type: replace-cross Abstract: This paper studies whether a lightweight supervised aggregator can combine diverse zero-shot large language model outputs into a stronger down

Learning-to-Explain through 20Q Gaming: An Explainable Recommender for Cybersecurity Education

SafetyDGX agent

arXiv:2604.26964v1 Announce Type: cross Abstract: The growing sophistication of contemporary cyber threats necessitates a more effective and adaptive approach to cybersecurity training. Intuitive and

Learning to Reason: Targeted Knowledge Discovery and Fuzzy Logic Update for Robust Image Recognition

ApplicationsDGX agent

arXiv:2604.27759v1 Announce Type: cross Abstract: Integrating domain knowledge into deep neural networks is a promising way to improve generalization. Existing methods either encode prior knowledge in

Learning to Spend: Model Predictive Control for Budgeting under Non-Stationary Returns

ResearchDGX agent

arXiv:2604.27186v1 Announce Type: cross Abstract: We study finite-horizon budget allocation as a closed-loop economic control problem and evaluate receding-horizon Model Predictive Control (MPC) relat

Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents

SafetyDGX agent

arXiv:2604.27283v1 Announce Type: cross Abstract: Large language model (LLM)-based coding agents increasingly rely on external memory to reuse prior debugging experience, repair traces, and repository

Lightweight Distillation of SAM 3 and DINOv3 for Edge-Deployable Individual-Level Livestock Monitoring and Longitudinal Visual Analytics

Model ReleasesDGX agent

arXiv:2604.27128v1 Announce Type: cross Abstract: Foundation-model pipelines for individual-level livestock monitoring -- combining open-vocabulary detection, promptable video segmentation, and self-s

LLM as Clinical Graph Structure Refiner: Enhancing Representation Learning in EEG Seizure Diagnosis

ResearchDGX agent

arXiv:2604.28178v1 Announce Type: new Abstract: Electroencephalogram (EEG) signals are vital for automated seizure detection, but their inherent noise makes robust representation learning challenging.

LLM Biases

SafetyDGX agent

arXiv:2604.26960v1 Announce Type: cross Abstract: Transformer-based agentic AI is rapidly being deployed on major platforms to help users shop, watch, and navigate content with less effort. While thes

LLMs as ASP Programmers: Self-Correction Enables Task-Agnostic Nonmonotonic Reasoning

ResearchDGX agent

arXiv:2604.27960v1 Announce Type: new Abstract: Recent large language models (LLMs) have achieved impressive reasoning milestones but continue to struggle with high computational costs, logical incons

Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation

Model ReleasesDGX agent

arXiv:2601.22228v2 Announce Type: replace-cross Abstract: We study whether vision-language models (VLMs) can solve relative camera pose estimation (RCPE) from image pairs, a direct test of multi-view

Machine Collective Intelligence for Explainable Scientific Discovery

AgentsDGX agent

arXiv:2604.27297v1 Announce Type: new Abstract: Deriving governing equations from empirical observations is a longstanding challenge in science. Although artificial intelligence (AI) has demonstrated

Making Conformal Predictors Robust in Healthcare Settings: a Case Study on EEG Classification

ApplicationsDGX agent

arXiv:2602.19483v2 Announce Type: replace-cross Abstract: Quantifying uncertainty in clinical predictions is critical for high-stakes diagnosis tasks. Conformal prediction offers a principled approach

Mapping how LLMs debate societal issues when shadowing human personality traits, sociodemographics and social media behavior

SafetyDGX agent

arXiv:2604.27624v1 Announce Type: cross Abstract: Large Language Models (LLMs) can strongly shape social discourse, yet datasets investigating how LLM outputs vary across controlled social and context

Mapping the Methodological Space of Classroom Interaction Research: Scale, Duration, and Modality in an Age of AI

TutorialsDGX agent

arXiv:2604.28098v1 Announce Type: new Abstract: Research on classroom interaction has long been divided between large-scale observation and in-depth ethnographic work. We propose a framework mapping t

Math Education Digital Shadows for facilitating learning with LLMs: Math performance, anxiety and confidence in simulated students and AIs

Model ReleasesDGX agent

arXiv:2604.27618v1 Announce Type: new Abstract: To enhance LLMs' impact on math education, we need data on their mathematical prowess and biases across prompts. To fill this gap, we introduce MEDS (Ma

MCPHunt: An Evaluation Framework for Cross-Boundary Data Propagation in Multi-Server MCP Agents

Model ReleasesDGX agent

arXiv:2604.27819v1 Announce Type: new Abstract: Multi-server MCP agents create an information-flow control problem: faithful tool composition can turn individually benign read/write permissions into c

Measurement Risk in Supervised Financial NLP: Rubric and Metric Sensitivity on JF-ICR

Model ReleasesDGX agent

arXiv:2604.27374v1 Announce Type: new Abstract: As LLMs become credible readers of earnings calls, investor-relations Q&A, guidance, and disclosure language, supervised financial NLP benchmarks increa

Mechanized Foundations of Structural Governance: Machine-Checked Proofs for Governed Intelligence

SafetyDGX agent

arXiv:2604.27289v1 Announce Type: new Abstract: We present five results in the theory of structural governance for cognitive workflow systems. Three are mechanized in Coq 8.19 using the Interaction Tr

METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution

SafetyDGX agent

arXiv:2604.27300v1 Announce Type: new Abstract: Metamaterial discovery seeks microstructured materials whose geometry induces targeted mechanical behavior. Existing inverse-design methods can efficien

MIFair: A Mutual-Information Framework for Intersectionality and Multiclass Fairness

SafetyDGX agent

arXiv:2604.28030v1 Announce Type: cross Abstract: Fairness in machine learning remains challenging due to its ethical complexity, the absence of a universal definition, and the need for context-specif

Mitigating Selection Bias in Large Language Models via Permutation-Aware GRPO

SafetyDGX agent

arXiv:2603.21016v2 Announce Type: replace-cross Abstract: Large language models (LLMs) used for multiple-choice and pairwise evaluation tasks often exhibit selection bias due to non-semantic factors l

Mixed Precision Training of Neural ODEs

ResearchDGX agent

arXiv:2510.23498v2 Announce Type: replace-cross Abstract: Exploiting low-precision computations has become a standard strategy in deep learning to address the growing computational costs imposed by ev

ML Code Smells: From Specification to Detection

ResearchDGX agent

arXiv:2509.20491v2 Announce Type: replace-cross Abstract: The rapid adoption of Artificial Intelligence (AI) is increasingly realised through Machine Learning (ML) pipelines that integrate data prepro

MM-StanceDet: Retrieval-Augmented Multi-modal Multi-agent Stance Detection

AgentsDGX agent

arXiv:2604.27934v1 Announce Type: new Abstract: Multimodal Stance Detection (MSD) is crucial for understanding public discourse, yet effectively fusing text and image, especially with conflicting sign

Modeling Clinical Concern Trajectories in Language Model Agents

AgentsDGX agent

arXiv:2604.27872v1 Announce Type: new Abstract: Large language model (LLM) agents deployed in clinical settings often exhibit abrupt, threshold-driven behavior, offering little visibility into accumul

Mull-Tokens: Modality-Agnostic Latent Thinking

ApplicationsDGX agent

arXiv:2512.10941v2 Announce Type: replace-cross Abstract: Reasoning goes beyond language; the real world requires reasoning about space, time, affordances, and much more that words alone cannot convey

Multibit neural inference in a N-ary crossbar architecture

ResearchDGX agent

arXiv:2604.26979v1 Announce Type: cross Abstract: In-memory computing (IMC) enables energy-efficient neural network inference by computing analog matrix-vector multiplications (MVM) in memory crossbar

NanoKnow: How to Know What Your Language Model Knows

Model ReleasesDGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

NeocorRAG: Less Irrelevant Information, More Explicit Evidence, and More Effective Recall via Evidence Chains

Model ReleasesDGX agent

arXiv:2604.27852v1 Announce Type: cross Abstract: Although precise recall is a core objective in Retrieval-Augmented Generation (RAG), a critical oversight persists in the field: improvements in retri

NORACL: Neurogenesis for Oracle-free Resource-Adaptive Continual Learning

TutorialsDGX agent

arXiv:2604.27031v1 Announce Type: cross Abstract: In a continual learning setting, we require a model to be plastic enough to learn a new task and stable enough to not disturb previously learned capab

Normativity and Productivism: Ableist Intelligence? A Degrowth Analysis of AI Sign Language Translation Tools for Deaf People

ResearchDGX agent

arXiv:2604.28125v1 Announce Type: new Abstract: Sign languages, of any geographical or accentual variation, understandably face continuous scrutiny under the ever present popularity of verbal dictatio

Not All Memories Age the Same: Autodiscovery of Adaptive Decay in Knowledge Graphs

ResearchDGX agent

arXiv:2604.26970v1 Announce Type: cross Abstract: Knowledge graphs used for retrieval treat all facts as equally current. Existing temporal approaches apply uniform decay, using a single forgetting cu

ObjectGraph: From Document Injection to Knowledge Traversal -- A Native File Format for the Agentic Era

Model ReleasesDGX agent

arXiv:2604.27820v1 Announce Type: new Abstract: Every document format in existence was designed for a human reader moving linearly through text. Autonomous LLM agents do not read - they retrieve. This

OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving

Local AiDGX agent

arXiv:2512.14044v3 Announce Type: replace-cross Abstract: The deployment of Vision-Language Models (VLMs) in safety-critical domains like autonomous driving (AD) is critically hindered by reliability

One Single Hub Text Breaks CLIP: Identifying Vulnerabilities in Cross-Modal Encoders via Hubness

ResearchDGX agent

arXiv:2604.27674v1 Announce Type: cross Abstract: The hubness problem, in which hub embeddings are close to many unrelated examples, occurs often in high-dimensional embedding spaces and may pose a pr

OpAgent: Operator Agent for Web Navigation

SafetyDGX agent

arXiv:2602.13559v2 Announce Type: replace Abstract: To fulfill user instructions, autonomous web agents must contend with the inherent complexity and volatile nature of real-world websites. Convention

OpenAI o1 System Card

SafetyDGX agent

arXiv:2412.16720v2 Announce Type: replace Abstract: The o1 model series is trained with large-scale reinforcement learning to reason using chain of thought. These advanced reasoning capabilities provi

OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research

Model ReleasesDGX agent

arXiv:2504.15564v3 Announce Type: replace-cross Abstract: Existing class-level code generation datasets are either synthetic (ClassEval: 100 classes) or insufficient in scale for modern training needs

Optimal Stop-Loss and Take-Profit Parameterization for Autonomous Trading Agent Swarm

AgentsDGX agent

arXiv:2604.27150v1 Announce Type: new Abstract: Autonomous crypto trading systems often spend most of their design effort on finding entries, while exits are left to fixed rules that are rarely tested

Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading

ResearchDGX agent

arXiv:2604.27637v1 Announce Type: new Abstract: Current Large Language Model (LLM) evaluation frameworks utilize the same static prompt template across all models under evaluation. This differs from t

OptimusKG: Unifying biomedical knowledge in a modern multimodal graph

AgentsDGX agent

arXiv:2604.27269v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) are widely used in the life sciences, yet many are derived from unstructured documents and therefore lack schema-level

OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment

Model ReleasesDGX agent

arXiv:2506.22500v2 Announce Type: replace-cross Abstract: Automated identification of surgical safety risks is critical for improving patient outcomes; however, Multimodal Large Language Models (MLLMs

ORFS-agent: Tool-Using Agents for Chip Design Optimization

Model ReleasesDGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

PALCAS: A Priority-Aware Intelligent Lane Change Advisory System for Autonomous Vehicles using Federated Reinforcement Learning

SafetyDGX agent

arXiv:2604.27118v1 Announce Type: cross Abstract: We present a priority-aware intelligent lane change advisory system based on multi-agent federated reinforcement learning, namely PALCAS, for autonomo

← Previous
1…289290291292293…354
Next →