AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
22 Apr 2026

Lost in the Prompt Order: Revealing the Limitations of Causal Attention in Language Models

ResearchDGX agent

arXiv:2601.14152v2 Announce Type: replace-cross Abstract: Large language models exhibit surprising sensitivity to the structure of the prompt, but the mechanisms underlying this sensitivity remain poo

Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning

SafetyDGX agent

arXiv:2604.18978v1 Announce Type: cross Abstract: Scaling critic capacity is a promising direction for enhancing off-policy reinforcement learning (RL). However, larger critics are prone to overfittin

LPO: Towards Accurate GUI Agent Interaction via Location Preference Optimization

AgentsDGX agent

arXiv:2506.09373v3 Announce Type: replace-cross Abstract: The advent of autonomous agents is transforming interactions with Graphical User Interfaces (GUIs) by employing natural language as a powerful


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LSTM-MAS: A Long Short-Term Memory Inspired Multi-Agent System for Long-Context Understanding

Model ReleasesDGX agent

arXiv:2601.11913v2 Announce Type: replace-cross Abstract: Effectively processing long contexts remains a fundamental yet unsolved challenge for large language models (LLMs). Existing single-LLM-based

Lyapunov-Certified Direct Switching Theory for Q-Learning

SafetyDGX agent

arXiv:2604.19569v1 Announce Type: cross Abstract: Q-learning is one of the most fundamental algorithms in reinforcement learning. We analyze constant-stepsize Q-learning through a direct stochastic sw

M^{2}GRPO: Mamba-based Multi-Agent Group Relative Policy Optimization for Biomimetic Underwater Robots Pursuit

SafetyDGX agent

arXiv:2604.19404v1 Announce Type: cross Abstract: Traditional policy learning methods in cooperative pursuit face fundamental challenges in biomimetic underwater robots, where long-horizon decision ma

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models

SafetyDGX agent

arXiv:2604.16755v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly integrated into daily life, in roles ranging from high-stakes decision support to companionship, un

MATA: Multi-Agent Framework for Reliable and Flexible Table Question Answering

AgentsDGX agent

arXiv:2602.09642v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have significantly improved table understanding tasks such as Table Question Answering (TableQ

Memory Assignment for Finite-Memory Strategies in Adversarial Patrolling Games

ResearchDGX agent

arXiv:2505.14137v2 Announce Type: replace Abstract: Adversarial Patrolling games form a subclass of Security games where a Defender moves between locations, guarding vulnerable targets. The main algor

Mesh Memory Protocol: Semantic Infrastructure for Multi-Agent LLM Systems

AgentsDGX agent

arXiv:2604.19540v1 Announce Type: cross Abstract: Teams of LLM agents increasingly collaborate on tasks spanning days or weeks: multi-day data-generation sprints where generator, reviewer, and auditor

Mind the (DH) Gap! A Contrast in Risky Choices Between Reasoning and Conversational LLMs

AgentsDGX agent

arXiv:2602.15173v2 Announce Type: replace Abstract: The use of large language models either as decision support systems, or in agentic workflows, is rapidly transforming the digital ecosystem. However

Modelling and Analysing Behaviours and Emotions via Complex User Interactions

TutorialsDGX agent

arXiv:1902.07683v1 Announce Type: cross Abstract: Over the past 15 years, the volume, richness and quality of data collected from the combined social networking platforms has increased beyond all expe

MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation

Model ReleasesDGX agent

arXiv:2604.18914v1 Announce Type: cross Abstract: While multilingual large language models (LLMs) perform well on high-level tasks like translation and question answering, their ability to handle gram

MRS: Multi-Resolution Skills for HRL Agents

SafetyDGX agent

arXiv:2505.21410v2 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) decomposes the policy into a manager and a worker, enabling long-horizon planning but introducing a perfor

Multi-Cycle Spatio-Temporal Adaptation in Human-Robot Teaming

TutorialsDGX agent

arXiv:2604.19670v1 Announce Type: cross Abstract: Effective human-robot teaming is crucial for the practical deployment of robots in human workspaces. However, optimizing joint human-robot plans remai

Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior

SafetyDGX agent

arXiv:2604.19102v1 Announce Type: cross Abstract: Learning diverse locomotion skills for humanoid robots in a unified reinforcement learning framework remains challenging due to the conflicting requir

Multi-Level Temporal Graph Networks with Local-Global Fusion for Industrial Fault Diagnosis

Local AiDGX agent

arXiv:2604.18765v1 Announce Type: cross Abstract: Fault detection and diagnosis are critical for the optimal and safe operation of industrial processes. The correlations among sensors often display no

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

SafetyDGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration

Model ReleasesDGX agent

arXiv:2604.19093v1 Announce Type: cross Abstract: Multi-modal test-time adaptation (TTA) enhances the resilience of benchmark multi-modal models against distribution shifts by leveraging the unlabeled

Multiclass Local Calibration with the Jensen-Shannon Distance

SafetyDGX agent

arXiv:2510.26566v2 Announce Type: replace-cross Abstract: Developing trustworthy Machine Learning (ML) models requires their predicted probabilities to be well-calibrated, meaning they should reflect

Multimodal Transformer for Sample-Aware Prediction of Metal-Organic Framework Properties

TutorialsDGX agent

arXiv:2604.19383v1 Announce Type: cross Abstract: Metal-organic frameworks (MOFs) are a major target of machine-learning-based property prediction, yet most models assume that a single framework repre

NeuroAI and Beyond: Bridging Between Advances in Neuroscience and ArtificialIntelligence

ResearchDGX agent

arXiv:2604.18637v1 Announce Type: cross Abstract: Neuroscience and Artificial Intelligence (AI) have made impressive progress in recent years but remain only loosely interconnected. Based on a worksho

Neuromorphic Continual Learning for Sequential Deployment of Nuclear Plant Monitoring Systems

SafetyDGX agent

arXiv:2604.18611v1 Announce Type: cross Abstract: Anomaly detection in nuclear industrial control systems (ICS) requires continuous, energy-efficient monitoring across multiple subsystems that are oft

Nexusformer: Nonlinear Attention Expansion for Stable and Inheritable Transformer Scaling

ResearchDGX agent

arXiv:2604.19147v1 Announce Type: cross Abstract: Scaling Transformers typically necessitates training larger models from scratch, as standard architectures struggle to expand without discarding learn

ODMA: On-Demand Memory Allocation Strategy for LLM Serving on LPDDR-Class Accelerators

Model ReleasesDGX agent

arXiv:2512.09427v5 Announce Type: replace-cross Abstract: Existing memory management techniques severely hinder efficient Large Language Model serving on accelerators constrained by poor random-access

OLLM: Options-based Large Language Models

Model ReleasesDGX agent

arXiv:2604.19087v1 Announce Type: new Abstract: We introduce Options LLM (OLLM), a simple, general method that replaces the single next-token prediction of standard LLMs with a extit{set of learned op

OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration

AgentsDGX agent

arXiv:2505.11765v3 Announce Type: replace-cross Abstract: Agents powered by advanced large language models (LLMs) have demonstrated impressive capabilities across diverse complex applications. Recentl

OmniGen2: Towards Instruction-Aligned Multimodal Generation

Model ReleasesDGX agent

arXiv:2506.18871v4 Announce Type: replace-cross Abstract: In this work, we introduce OmniGen2, a versatile and open-source generative model designed to provide a unified solution for diverse generatio

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens

Model ReleasesDGX agent

arXiv:2604.18827v1 Announce Type: cross Abstract: Scaling data and artificial neural networks has transformed AI, driving breakthroughs in language and vision. Whether similar principles apply to mode

On Accelerating Grounded Code Development for Research

ResearchDGX agent

arXiv:2604.19022v1 Announce Type: new Abstract: A major challenge for niche scientific and technical domains in leveraging coding agents is the lack of access to up-to-date, domain- specific knowledge

On Solving the Multiple Variable Gapped Longest Common Subsequence Problem

ResearchDGX agent

arXiv:2604.18645v1 Announce Type: new Abstract: This paper addresses the Variable Gapped Longest Common Subsequence (VGLCS) problem, a generalization of the classical LCS problem involving flexible ga

On the Spatiotemporal Dynamics of Generalization in Neural Networks

Local AiDGX agent

arXiv:2602.01651v2 Announce Type: replace-cross Abstract: Why do neural networks fail to generalize addition from 16-digit to 32-digit numbers, while a child who learns the rule can apply it to arbitr

One Step Forward and K Steps Back: Better Reasoning with Denoising Recursion Models

Model ReleasesDGX agent

arXiv:2604.18839v1 Announce Type: cross Abstract: Looped transformers scale computational depth without increasing parameter count by repeatedly applying a shared transformer block and can be used for

Ontology-Constrained Neural Reasoning in Enterprise Agentic Systems: A Neurosymbolic Architecture for Domain-Grounded AI Agents

Local AiDGX agent

arXiv:2604.00555v2 Announce Type: replace Abstract: Enterprise adoption of Large Language Models (LLMs) is constrained by hallucination, domain drift, and the inability to enforce regulatory complianc

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models

Model ReleasesDGX agent

arXiv:2509.15435v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) exhibit strong multimodal capabilities but remain vulnerable to hallucinations from intrinsic errors and

Owner-Harm: A Missing Threat Model for AI Agent Safety

Model ReleasesDGX agent

arXiv:2604.18658v1 Announce Type: cross Abstract: Existing AI agent safety benchmarks focus on generic criminal harm (cybercrime, harassment, weapon synthesis), leaving a systematic blind spot for a d

Personalized Benchmarking: Evaluating LLMs by Individual Preferences

SafetyDGX agent

arXiv:2604.18943v1 Announce Type: new Abstract: With the rise in capabilities of large language models (LLMs) and their deployment in real-world tasks, evaluating LLM alignment with human preferences

PhysMem: Scaling Test-time Physical Memory for Robot Manipulation

TutorialsDGX agent

arXiv:2602.20323v5 Announce Type: replace-cross Abstract: Reliable object manipulation requires understanding physical properties that vary across objects and environments. Vision-language model (VLM)

PLaMo 2.1-VL Technical Report

AgentsDGX agent

arXiv:2604.19324v1 Announce Type: cross Abstract: We introduce PLaMo 2.1-VL, a lightweight Vision Language Model (VLM) for autonomous devices, available in 8B and 2B variants and designed for local an

Plausible Reasoning and First-Order Plausible Logic

ResearchDGX agent

arXiv:2604.19036v1 Announce Type: new Abstract: Defeasible statements are statements that are likely, or probable, or usually true, but may occasionally be false. Plausible reasoning makes conclusions

Position: No Retroactive Cure for Infringement during Training

ApplicationsDGX agent

arXiv:2604.18649v1 Announce Type: cross Abstract: As generative AI faces intensifying legal challenges, the machine learning community has increasingly relied on post-hoc mitigation -- especially mach

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

SafetyDGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety

SafetyDGX agent

arXiv:2604.19083v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in cross-modal understanding and generation, yet their deployment is threate

Prompt to Pwn: Automated Exploit Generation for Smart Contracts

ApplicationsDGX agent

arXiv:2508.01371v3 Announce Type: replace-cross Abstract: Smart contracts are important for digital finance, yet they are hard to patch once deployed. Prior work has mainly explored LLMs for smart con

Protecting Bystander Privacy via Selective Hearing in Audio LLMs

Model ReleasesDGX agent

arXiv:2512.06380v3 Announce Type: replace-cross Abstract: Audio Large language models (LLMs) are increasingly deployed in the real world, where they inevitably capture speech from unintended nearby by

PuzzleWorld: A Benchmark for Multimodal, Open-Ended Reasoning in Puzzlehunts

Model ReleasesDGX agent

arXiv:2506.06211v2 Announce Type: replace-cross Abstract: Puzzlehunts are a genre of complex, multi-step puzzles lacking well-defined problem definitions. In contrast to conventional reasoning benchma

QSLM: A Performance- and Memory-aware Quantization Framework with Tiered Search Strategy for Spike-driven Language Models

ResearchDGX agent

arXiv:2601.00679v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been emerging as prominent AI models for solving many natural language tasks due to their high performance (

Quantum inspired qubit qutrit neural networks for real time financial forecasting

ResearchDGX agent

arXiv:2604.18838v1 Announce Type: new Abstract: This research investigates the performance and efficacy of machine learning models in stock prediction, comparing Artificial Neural Networks (ANNs), Qua

Quantum spatial best-arm identification via quantum walks

ResearchDGX agent

arXiv:2509.05890v3 Announce Type: replace-cross Abstract: Quantum reinforcement learning has emerged as a framework combining quantum computation with sequential decision-making, and applications to t

R^2-dLLM: Accelerating Diffusion Large Language Models via Spatio-Temporal Redundancy Reduction

Local AiDGX agent

arXiv:2604.18995v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to autoregressive generation by enabling parallel token prediction. Ho

RARE: Redundancy-Aware Retrieval Evaluation Framework for High-Similarity Corpora

Model ReleasesDGX agent

arXiv:2604.19047v1 Announce Type: cross Abstract: Existing QA benchmarks typically assume distinct documents with minimal overlap, yet real-world retrieval-augmented generation (RAG) systems operate o

RDP LoRA: Geometry-Driven Identification for Parameter-Efficient Adaptation in Large Language Models

Model ReleasesDGX agent

arXiv:2604.19321v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) remains structurally uncertain despite parameter-efficient methods such as Low-Rank Adaptation (LoRA), as the

Reasoning-Aware AIGC Detection via Alignment and Reinforcement

SafetyDGX agent

arXiv:2604.19172v1 Announce Type: new Abstract: The rapid advancement and widespread adoption of Large Language Models (LLMs) have elevated the need for reliable AI-generated content (AIGC) detection,

Reasoning Over Space: Enabling Geographic Reasoning for LLM-Based Generative Next POI Recommendation

Local AiDGX agent

arXiv:2601.04562v2 Announce Type: replace Abstract: Generative recommendation with large language models (LLMs) reframes prediction as sequence generation, yet existing LLM-based recommenders remain l

Reasoning Structure Matters for Safety Alignment of Reasoning Models

SafetyDGX agent

arXiv:2604.18946v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve strong performance on complex reasoning tasks but often generate harmful responses to malicious user queries. This

Reduced-Order Surrogates for Forced Flexible Mesh Coastal-Ocean Models

ApplicationsDGX agent

arXiv:2602.05416v2 Announce Type: replace-cross Abstract: While proper orthogonal decomposition (POD)-based surrogates are widely explored for hydrodynamic applications, the use of Koopman autoencoder

Reducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularization

ResearchDGX agent

arXiv:2604.19079v1 Announce Type: cross Abstract: Unification of automatic speech recognition (ASR) systems reduces development and maintenance costs, but training a single model to perform well in bo

ReefNet: A Large-Scale Dataset and Benchmark for Fine-Grained Coral Reef Recognition

Model ReleasesDGX agent

arXiv:2510.16822v3 Announce Type: replace-cross Abstract: Coral reefs are rapidly declining under anthropogenic pressures (e.g., climate change), creating an urgent need for scalable and automated mon

Refute-or-Promote: An Adversarial Stage-Gated Multi-Agent Review Methodology for High-Precision LLM-Assisted Defect Discovery

AgentsDGX agent

arXiv:2604.19049v1 Announce Type: cross Abstract: LLM-assisted defect discovery has a precision crisis: plausible-but-wrong reports overwhelm maintainers and degrade credibility for real findings. We

Regulating Artificial Intimacy: From Locks and Blocks to Relational Accountability

SafetyDGX agent

arXiv:2604.18893v1 Announce Type: cross Abstract: A series of high-profile tragedies involving companion chatbots has triggered an unusually rapid regulatory response. Several jurisdictions, including

← Previous
1…318319320321322…354
Next →