AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
19 May 2026

MemRepair: Hierarchical Memory for Agentic Repository-Level Vulnerability Repair

AgentsDGX agent

arXiv:2605.17444v1 Announce Type: cross Abstract: Modern software ecosystems face a rapidly growing number of disclosed vulnerabilities, increasing the need for automated repair techniques that can op

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

Model ReleasesDGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks

ResearchDGX agent

arXiv:2605.16864v1 Announce Type: cross Abstract: Although large-scale visual foundation models (VFMs) achieve remarkable performance in semantic understanding, they still underperform in instance-awa


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MHMamba: Multi-Head Mamba for 3D Brain Tumor Segmentation

ResearchDGX agent

arXiv:2605.16464v1 Announce Type: cross Abstract: Brain tumors exhibit high heterogeneity in morphology and multimodal contrast, making manual slice-by-slice de lineation time-consuming and experience

Minor First, Major Last: A Depth-Induced Implicit Bias of Sharpness-Aware Minimization

SafetyDGX agent

arXiv:2603.08290v2 Announce Type: replace-cross Abstract: We study the implicit bias of Sharpness-Aware Minimization (SAM) when training L-layer linear diagonal networks on linearly separable binary c

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

Model ReleasesDGX agent

arXiv:2601.08118v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as human simulators, both for evaluating conversational systems and for generating fine-tuning da

Missing-Modality-Aware Graph Neural Network for Cancer Classification

ApplicationsDGX agent

arXiv:2506.22901v2 Announce Type: replace-cross Abstract: A key challenge in learning from multimodal biological data is missing modalities, where data from one or more modalities are absent for some

Mitigating Conversational Inertia in Multi-Turn Agents

SafetyDGX agent

arXiv:2602.03664v3 Announce Type: replace Abstract: Large language models excel as few-shot learners when provided with appropriate demonstrations, yet this strength becomes problematic in multiturn a

Mixing Times of Glauber Dynamics on Masked Language Models

ApplicationsDGX agent

arXiv:2605.16378v1 Announce Type: cross Abstract: Masked language models (MLMs) define local conditional distributions over tokens but do not, in general, correspond to any consistent joint distributi

Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource

Model ReleasesDGX agent

arXiv:2506.12119v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) language models dramatically expand model capacity and achieve remarkable performance without increasing per-token co

Modality vs. Morphology: A Framework for Time Series Classification for Biological Signals

ResearchDGX agent

arXiv:2605.18483v1 Announce Type: cross Abstract: Time series classification (TSC) of biological signals has progressed from handcrafted, modality-specific approaches to deep architectures capable of

Modelling Customer Trajectories with Reinforcement Learning for Practical Retail Insights

AgentsDGX agent

arXiv:2605.18449v1 Announce Type: cross Abstract: Understanding customer movement within retail spaces is essential for optimizing store layouts. Real-world trajectory data can provide highly accurate

MoleCode unlocks structural intelligence in large language models

ResearchDGX agent

arXiv:2605.16480v1 Announce Type: cross Abstract: Molecules are graphs, but large language models~(LLMs) are usually asked to reason about them through linear strings. The most popular molecular repre

MR-SLAM: Immersive Spatial Supervision for Multi-Robot Mapping via Mixed Reality

ApplicationsDGX agent

arXiv:2605.16432v1 Announce Type: cross Abstract: Operating a multi-robot fleet for simultaneous localization and mapping (SLAM) in applications such as building inspection or warehouse-aisle monitori

Multi-agent AI systems outperform human teams in creativity

Local AiDGX agent

arXiv:2605.17885v1 Announce Type: cross Abstract: Although artificial intelligence (AI) now matches or exceeds human performance across numerous cognitive tasks, creativity remains a highly contested

Multi-Object Tracking Consistently Improves Wildlife Inference

ApplicationsDGX agent

arXiv:2605.16672v1 Announce Type: cross Abstract: Camera traps have become a common tool for wildlife monitoring efforts in ecological research and biodiversity conservation. Wildlife classification m

Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework

AgentsDGX agent

arXiv:2605.16821v1 Announce Type: new Abstract: The rapid evolution of Large Language Model (LLM) agents has produced diverse interaction paradigms, yet few production systems integrate multiple parad

Multi-Party Multi-Objective Optimization as Consensus Search: Runtime Analysis of Cross-Party Recombination

Model ReleasesDGX agent

arXiv:2605.17454v1 Announce Type: new Abstract: Multi-party multi-objective optimization problems (MPMOPs) require consensus among autonomous decision makers and therefore differ from flattened many-o

Multi-task learning on partially labeled datasets via invariant/equivariant semi-supervised learning

ResearchDGX agent

arXiv:2605.17624v1 Announce Type: cross Abstract: We investigate the potential of invariant and equivariant semi-supervised learning for addressing the challenges of training multi-task models on part

Multilingual jailbreaking of LLMs using low-resource languages

Model ReleasesDGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

Model ReleasesDGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

MusicSynth: An Automated Pipeline for Generating Violin Fingerboard Animations from Sheet Music Using Optical Music Recognition

TutorialsDGX agent

arXiv:2605.17181v1 Announce Type: cross Abstract: Learning the violin is harder than it looks. Unlike piano keys or guitar frets, the violin neck has no markings at all, so a beginner cannot tell by l

Mutual Enhancement Between Global Tokens and Patch Tokens: From Theory to Practice

ResearchDGX agent

arXiv:2605.16384v1 Announce Type: cross Abstract: Accurate and effective discrete image tokenization is crucial for long image sequence processing. However, current methods rigidly compress all conten

Natural-Language Agent Harnesses

SafetyDGX agent

arXiv:2603.25723v2 Announce Type: replace-cross Abstract: Agent performance is strongly shaped by the surrounding harness: the external execution system around a model that organizes a task run. Yet t

Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior

ResearchDGX agent

arXiv:2502.20349v3 Announce Type: replace-cross Abstract: How can cognitive science build generalizable theories that span the full scope of natural situations and behaviors? We argue that progress in

Needles in the Landscape: Semi-Supervised Pseudolabeling for Archaeological Site Discovery under Label Scarcity

ResearchDGX agent

arXiv:2510.16814v2 Announce Type: replace-cross Abstract: Archaeological predictive modelling estimates where undiscovered sites are likely to occur by combining known locations with environmental, cu

Nested Spatio-Temporal Time Series Forecasting

ApplicationsDGX agent

arXiv:2605.16447v1 Announce Type: cross Abstract: Spatiotemporal forecasting is critical for real-world applications like traffic management, yet capturing reliable interactions remains challenging un

Neural Visual Decoding via Cognitive guided Adaptive Blurring and Information Constrained Alignment

SafetyDGX agent

arXiv:2605.16418v1 Announce Type: cross Abstract: EEG-based visual decoding aims to establish a mapping between neural signals and visual semantics. However, it remains constrained by the dual challen

NeuroMAS: Multi-Agent Systems as Neural Networks with Joint Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.16757v1 Announce Type: new Abstract: Multi-agent language systems are often built as hand-designed workflows, where agents are assigned semantic roles and communication protocols are specif

NeuroRVQ: Multi-Scale Biosignal Tokenization for Generative Foundation Models

ResearchDGX agent

arXiv:2510.13068v4 Announce Type: replace-cross Abstract: Biosignals such as electroencephalography (EEG), electrocardiography (ECG), and electromyography (EMG) encode physiological activity across mu

NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents

Model ReleasesDGX agent

arXiv:2605.17596v1 Announce Type: new Abstract: We present NeuSymMS, an adaptive memory system that enables large language model (LLM) agents to learn, remember, and reason about users across sessions

New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity Contradictions

ResearchDGX agent

arXiv:2605.18035v1 Announce Type: new Abstract: Hard-thresholding is an important type of algorithm in machine learning that is used to solve ell_0 constrained optimization problems. However, the true

New Wide-Net-Casting Jailbreak Attacks Risk Large Models

SafetyDGX agent

arXiv:2605.17128v1 Announce Type: cross Abstract: Jailbreak attacks on large models have drawn growing attention due to their close ties to societal safety. This work identifies a practical yet unexpl

NGM: A Plug-and-Play Training-Free Memory Module for LLMs

ResearchDGX agent

arXiv:2605.16893v1 Announce Type: new Abstract: Recent studies introduce conditional memory modules that decouple knowledge storage from neural computation, enabling more direct knowledge access. Comp

No Plan, Yet Human: A Reactive Robotics Model Predicts Human Planning Failures on a Clinical Task

ResearchDGX agent

arXiv:2605.16514v1 Announce Type: cross Abstract: Understanding why some sequential planning problems are harder than others requires models that go beyond average performance. They should capture the

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

Model ReleasesDGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

Observation-Aligned Mask Priors for Learning Physical Dynamics from Authentic Occlusions

TutorialsDGX agent

arXiv:2605.16818v1 Announce Type: cross Abstract: Learning physical dynamics directly from incomplete observations is challenging because authentic occlusions are structured, sample-dependent, and oft

OCCAM: Open-set Causal Concept explAnation and Ontology induction for black-box vision Models

Local AiDGX agent

arXiv:2605.18481v1 Announce Type: new Abstract: Interpreting the decisions of deep image classifiers remains challenging, particularly in black-box settings where model internals are inaccessible. We

Old Habits Die Hard: How Conversational History Geometrically Traps LLMs

SafetyDGX agent

arXiv:2603.03308v2 Announce Type: replace-cross Abstract: How does the conversational past of large language models (LLMs) influence their future performance? Recent work suggests that LLMs are affect

oldsymbol{f}-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control

SafetyDGX agent

arXiv:2605.17862v1 Announce Type: cross Abstract: Scaling on-policy distillation (OPD) for large language models (LLMs) confronts a fundamental tension: asynchronous execution is necessary for system

OmniCode: A Benchmark for Evaluating Software Engineering Agents

Model ReleasesDGX agent

arXiv:2602.02262v3 Announce Type: replace-cross Abstract: LLM-powered coding agents are redefining how real-world software is developed. To drive the research towards better coding agents, we require

OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics

Model ReleasesDGX agent

arXiv:2605.16962v1 Announce Type: cross Abstract: Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the

On Safer Reinforcement Learning for Sedation and Analgesia in Intensive Care

SafetyDGX agent

arXiv:2601.23154v2 Announce Type: replace-cross Abstract: Pain management in intensive care usually involves complex trade-offs, since both inadequate and excessive treatment can compromise patient sa

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression

SafetyDGX agent

arXiv:2601.21531v2 Announce Type: replace-cross Abstract: Visual token compression is widely used to accelerate large vision-language models (LVLMs) by pruning or merging visual tokens, yet its advers

One Model to Translate Them All: Universal Any-to-Any Translation for Heterogeneous Collaborative Perception

AgentsDGX agent

arXiv:2605.17907v1 Announce Type: cross Abstract: By sharing intermediate features, collaborative perception extends each agent's sensing beyond standalone limits, but real-world feature modality hete

One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer

Model ReleasesDGX agent

arXiv:2605.17811v1 Announce Type: cross Abstract: Can a shared-weight recurrent Transformer develop distinct internal roles without being partitioned into separate modules? We study this in Asymmetric

Online Algorithms with Unreliable Guidance

TutorialsDGX agent

arXiv:2602.20706v2 Announce Type: replace Abstract: This paper introduces online algorithms with unreliable guidance (OAG), a model for ML-augmented online decision-making that cleanly separates the p

OpenJarvis: Personal AI, On Personal Devices

Model ReleasesDGX agent

arXiv:2605.17172v1 Announce Type: cross Abstract: Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local

OPERA: A Reinforcement Learning--Enhanced Orchestrated Planner-Executor Architecture for Reasoning-Oriented Multi-Hop Retrieval

SafetyDGX agent

arXiv:2508.16438v4 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) and dense retrievers have driven significant progress in retrieval-augmented generation (RAG).

OProver: A Unified Framework for Agentic Formal Theorem Proving

AgentsDGX agent

arXiv:2605.17283v1 Announce Type: cross Abstract: Recent progress in formal theorem proving has benefited from large-scale proof generation and verifier-aware training, but agentic proving is rarely i

Optimal Knock-Pick Planning for Tightly Packed Tabletop Blocks With Parallel Grippers

ResearchDGX agent

arXiv:2605.17800v1 Announce Type: cross Abstract: Rearranging densely packed tabletop objects is challenging when parallel-gripper picks are infeasible without sufficient clearance around an object. T

Optimising CSRNet with parameter-free attention mechanisms for crowd counting in public transport

Model ReleasesDGX agent

arXiv:2605.18349v1 Announce Type: cross Abstract: Occupancy estimation and crowd counting are critical tasks in designing smart and efficient public transport vehicles. Given that public transport loa

Ordinal Adaptive Correction: A Data-Centric Approach to Ordinal Image Classification with Noisy Labels

Model ReleasesDGX agent

arXiv:2509.02351v3 Announce Type: replace-cross Abstract: Labeled data is a fundamental component in training supervised deep learning models for computer vision tasks. However, the labeling process,

Orthologic for SAT Solving

ResearchDGX agent

arXiv:2605.16421v1 Announce Type: cross Abstract: We present a new algorithm for deciding formula entailment in orthologic (a sound approximation of classical logic) that avoids the costly preprocessi

OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization

ResearchDGX agent

arXiv:2605.17757v1 Announce Type: cross Abstract: INT2 KV-cache quantization is attractive for long-context LLM serving, but it remains difficult to make both accurate and deployable. Simple rotations

OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents

Model ReleasesDGX agent

arXiv:2506.16042v2 Announce Type: replace Abstract: Generative AI is being leveraged to solve a variety of computer-use tasks involving desktop applications. State-of-the-art systems have focused sole

Overcoming the Intrinsic Performance Limitations of MEMS IMU via Diffusion-Based Generative Learning

Local AiDGX agent

arXiv:2605.16391v1 Announce Type: cross Abstract: Inertial measurement units (IMUs) are fundamental sensing components in multi-source integrated navigation systems, and their performance directly det

Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks

Model ReleasesDGX agent

arXiv:2605.18583v1 Announce Type: cross Abstract: Coding agents now run autonomously with shell, file, and network privileges. When a user issues a benign request, the agent sometimes does more than a

OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism

Local AiDGX agent

arXiv:2603.14371v2 Announce Type: replace-cross Abstract: Embodied AI agents increasingly require parallel execution of multiple tasks, such as manipulation, conversation, and memory construction, fro

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization

SafetyDGX agent

arXiv:2605.17877v1 Announce Type: new Abstract: A significant hurdle for current LLMs is the execution of complex, multi-stage tasks. Group Relative Policy Optimization (GRPO) has been emerging as a l

← Previous
1…239240241242243…358
Next →