AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
11 May 2026

Dynamic one-time delivery of critical data by small and sparse UAV swarms: a model problem for MARL scaling studies

SafetyDGX agent

arXiv:2512.09682v2 Announce Type: replace-cross Abstract: This work studies the application of Multi-Agent Reinforcement Learning (MARL) to decentralized control of unmanned aerial vehicles to relay a

Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies

SafetyDGX agent

arXiv:2603.00041v2 Announce Type: replace-cross Abstract: Causal machine learning (ML) recovers graphical structures that inform us about potential cause-and-effect relationships. Most progress has fo

Edge Deep Learning in Computer Vision and Medical Diagnostics: A Comprehensive Survey

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.06714v1 Announce Type: cross Abstract: Edge deep learning, a paradigm change reconciling edge computing and deep learning, facilitates real-time decision making attuned to environmental fac

Efficient Data Selection for Multimodal Models via Incremental Optimization Utility

Model ReleasesDGX agent

arXiv:2605.07488v1 Announce Type: new Abstract: The scaling of Large Multimodal Models (LMMs) is constrained by the quality-quantity trade-off inherent in synthetic data. Previous approaches, such as

EGA: Adapting Frozen Encoders for Vector Search with Bounded Out-of-Distribution Degradation

SafetyDGX agent

arXiv:2605.05674v2 Announce Type: replace-cross Abstract: Vector search systems built on frozen vision encoders face queries from unseen classes at deployment, yet existing adapter training collapses

EgoPro-Bench: Benchmarking Personalized Proactive Interaction in Egocentric Video Streams

Model ReleasesDGX agent

arXiv:2605.07299v1 Announce Type: cross Abstract: Existing Multimodal Large Language Models (MLLMs) remain primarily reactive, failing to continuously perceive environments or proactively assist users

EmambaIR: Efficient Visual State Space Model for Event-guided Image Reconstruction

TutorialsDGX agent

arXiv:2605.08073v1 Announce Type: cross Abstract: Recent event-based image reconstruction methods predominantly rely on Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs) to process c

Emergent social transmission of model-based representations without inference

SafetyDGX agent

arXiv:2604.05777v2 Announce Type: replace Abstract: How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to r

Enabling Unsupervised Training of Deep EEG Denoisers With Intelligent Partitioning

ResearchDGX agent

arXiv:2605.06724v1 Announce Type: cross Abstract: Denoising wearable electroencephalogram (EEG) is inherently challenging since neural activity is not only subtle but also inseparable from spectrally

End-to-end PDDL Planning with Hardcoded and Dynamic Agents

Model ReleasesDGX agent

arXiv:2512.09629v2 Announce Type: replace Abstract: We present an end-to-end framework for planning supported by verifiers. An orchestrator receives a human specification written in natural language a

Ensemble Distributionally Robust Bayesian Optimisation

ResearchDGX agent

arXiv:2605.07565v1 Announce Type: cross Abstract: We study zeroth-order optimisation under context distributional uncertainty, a setting commonly tackled using Bayesian optimisation (BO). A prevailing

Ensemble Learning for Healthcare: A Comparative Analysis of Hybrid Voting and Ensemble Stacking in Obesity Risk Prediction

ApplicationsDGX agent

arXiv:2509.02826v2 Announce Type: replace-cross Abstract: Obesity is a critical global health issue driven by dietary, physiological, and environmental factors, and is strongly associated with chronic

Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co

EnvSimBench: A Benchmark for Evaluating and Improving LLM-Based Environment Simulation

Model ReleasesDGX agent

arXiv:2605.07247v1 Announce Type: new Abstract: Scalable AI agents training relies on interactive environments that faithfully simulate the consequences of agent actions. Manually crafted environments

Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation

SafetyDGX agent

arXiv:2605.07381v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models offer broad general capabilities, deploying them on specific hardware requires real-world adaptation to brid

ESSAM: A Novel Competitive Evolution Strategies Approach to Reinforcement Learning for Memory Efficient LLMs Fine-Tuning

Model ReleasesDGX agent

arXiv:2602.01003v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a key training step for improving mathematical reasoning in large language models (LLMs), but it often

EULER-ADAS: Energy-Efficient & SIMD-Unified Logarithmic-Posit Engine for Precision-Reconfigurable Approximate ADAS Acceleration

ResearchDGX agent

arXiv:2605.06875v1 Announce Type: cross Abstract: Advanced driver-assistance systems (ADAS) require neural compute engines that deliver low-latency inference under strict power and area constraints. P

Evaluating Large Language Models in Scientific Discovery

Model ReleasesDGX agent

arXiv:2512.15567v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet prevailing science benchmarks probe decontextualized knowledge and

Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs

Model ReleasesDGX agent

arXiv:2605.06669v1 Announce Type: cross Abstract: Educational LLM tutors face a core AI alignment challenge: they must follow user intent while preserving pedagogical constraints and safety policies.

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

SafetyDGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

Model ReleasesDGX agent

arXiv:2603.06859v2 Announce Type: replace-cross Abstract: Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts t

Exact Regular-Constrained Variable-Order Markov Generation via Sparse Context-State Belief Propagation

ResearchDGX agent

arXiv:2605.07839v1 Announce Type: new Abstract: Variable-order Markov models generate sequences over a finite alphabet by conditioning each symbol on the longest available suffix of the generated hist

Excluding the Target Domain Improves Extrapolation: Deconfounded Hierarchical Physics Constraints

Model ReleasesDGX agent

arXiv:2605.07485v1 Announce Type: cross Abstract: Extrapolation to out-of-distribution conditions is a fundamental challenge for physics-constrained deep generative models. Existing methods apply phys

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models

SafetyDGX agent

arXiv:2605.07244v1 Announce Type: cross Abstract: We introduce Mutual Reinforcement Learning, a framework for concurrent RL post-training in which heterogeneous LLM policies exchange typed experience

Exploring the non-convexity in machine learning using quantum-inspired optimization

ResearchDGX agent

arXiv:2605.07947v1 Announce Type: cross Abstract: The escalating complexity of modern machine learning necessitates solving challenging non-convex optimization problems, particularly in high-dimension

Exposing and Mitigating Temporal Attack in Deepfake Video Detection

ResearchDGX agent

arXiv:2605.07398v1 Announce Type: cross Abstract: While spatiotemporal deepfake detectors achieve high AUC, our experiments reveal their susceptibility to evasion attacks. These models tend to overfit

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning

ResearchDGX agent

arXiv:2605.06840v1 Announce Type: new Abstract: Large language models (LLMs), especially reasoning models, generate extended chain-of-thought (CoT) reasoning that often contains explicit deliberation

f-Divergence Regularized RLHF: Two Tales of Sampling and Unified Analyses

SafetyDGX agent

arXiv:2605.06977v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a cornerstone technique for post-training large language models. While most existing appr

Factored Classifier-Free Guidance

ResearchDGX agent

arXiv:2506.14399v5 Announce Type: replace-cross Abstract: Counterfactual generation aims to simulate realistic hypothetical outcomes under causal interventions. Diffusion models have emerged as a powe

FactoryBench: Evaluating Industrial Machine Understanding

Model ReleasesDGX agent

arXiv:2605.07675v1 Announce Type: new Abstract: We introduce FactoryBench, a benchmark for evaluating time-series models and LLMs on machine understanding over industrial robotic telemetry. Q&A pairs

Fast and Effective Redistricting Optimization via Composite-Move Tabu Search

ApplicationsDGX agent

arXiv:2605.06682v1 Announce Type: new Abstract: Spatial redistricting is a practical combinatorial optimization problem that demands high-quality solutions, rapid turnaround, and flexibility to accomm

Fast Byte Latent Transformer

Model ReleasesDGX agent

arXiv:2605.08044v1 Announce Type: cross Abstract: Recent byte-level language models (LMs) match the performance of token-level models without relying on subword vocabularies, yet their utility is limi

Federated Spatiotemporal Graph Learning for Passive Attack Detection in Smart Grids

Local AiDGX agent

arXiv:2510.02371v2 Announce Type: replace-cross Abstract: Smart grids are exposed to passive eavesdropping, where attackers listen silently to communication links. Although no data is actively altered

Finite-Time Analysis of MCTS in Continuous POMDP Planning

ResearchDGX agent

arXiv:2605.07703v1 Announce Type: new Abstract: This paper presents a finite-time analysis for Monte Carlo Tree Search (MCTS) in Partially Observable Markov Decision Processes (POMDPs), with probabili

FiSMiness: A Finite State Machine Based Paradigm for Emotional Support Conversations

ResearchDGX agent

arXiv:2504.11837v2 Announce Type: replace-cross Abstract: Emotional support conversation (ESC) aims to alleviate the emotional distress of individuals through effective conversations. Although large l

FlashMol: High-Quality Molecule Generation in as Few as Four Steps

ResearchDGX agent

arXiv:2605.07020v1 Announce Type: cross Abstract: Generating chemically valid 3D molecular conformations is critical for computational drug discovery. Classical diffusion-based models like GeoLDM perf

Flat Channels to Infinity in Neural Loss Landscapes

Model ReleasesDGX agent

arXiv:2506.14951v4 Announce Type: replace-cross Abstract: The loss landscapes of neural networks contain minima and saddle points that may be connected in flat regions or appear in isolation. We ident

Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

SafetyDGX agent

arXiv:2602.09782v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a critical method for enhancing the reasoning capabilities of Large Langu

Flow-OPD: On-Policy Distillation for Flow Matching Models

SafetyDGX agent

arXiv:2605.08063v1 Announce Type: cross Abstract: Existing Flow Matching (FM) text-to-image models suffer from two critical bottlenecks under multi-task alignment: the reward sparsity induced by scala

ForgeVLA: Federated Vision-Language-Action Learning without Language Annotations

TutorialsDGX agent

arXiv:2605.07474v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models hold great promise for general-purpose robotic intelligence, yet scaling up such models is severely bottlenecked b

Frequency-Aware Model Parameter Explorer: A new attribution method for improving explainability

Model ReleasesDGX agent

arXiv:2510.03245v2 Announce Type: replace-cross Abstract: State-of-the-art attribution methods rely on adversarial sample generation that applies an all-pass filter across the frequency spectrum, disc

From Assistance to Agency: Rethinking Autonomy and Control in CI/CD Pipelines

SafetyDGX agent

arXiv:2605.07062v1 Announce Type: cross Abstract: AI agents are assuming active roles in Continuous Integration and Continuous Deployment (CI/CD) workflows, yet the research community lacks a shared v

From Clouds to Hallucinations: Atmospheric Retrieval Hijacking in Remote Sensing Vision-Language RAG

Model ReleasesDGX agent

arXiv:2605.07273v1 Announce Type: cross Abstract: Multimodal RAG systems increasingly rely on vision-language retrievers to ground visual queries in external textual evidence. Existing adversarial stu

From Feasible to Practical: Pareto-Optimal Synthesis Planning

ApplicationsDGX agent

arXiv:2605.07521v1 Announce Type: new Abstract: Current computer-aided synthesis planning (CASP) methods often treat retrosynthesis as solved once a single feasible route is identified, focusing prima

From Pixels to Prompts: Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.07544v1 Announce Type: new Abstract: When you read a paper about a new Vision-Language Model today, it can be easy to forget how strange this idea would have sounded not so long ago. Teachi

From Specification to Deployment: Empirical Evidence from a W3C VC + DID Trust Infrastructure for Autonomous Agents

SafetyDGX agent

arXiv:2605.06738v1 Announce Type: cross Abstract: Autonomous AI agents now transact at production scale -- 69,000 bots executing 165 million transactions across 50 million USDC in cumulative volume on

From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms

AgentsDGX agent

arXiv:2605.06716v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents have fundamentally reshaped artificial intelligence by integrating external tools and planning capabilities. Whi

From Surface Learning to Deep Understanding: A Grounded AI Tutoring System for Moodle

ApplicationsDGX agent

arXiv:2605.06963v1 Announce Type: cross Abstract: This demo paper describes the development of the AI Teaching & Learning Assistant, a modular Moodle plugin that leverages Retrieval-Augmented Generati

From Time Series Analysis to Question Answering: A Survey in the LLM Era

SafetyDGX agent

arXiv:2506.11512v2 Announce Type: replace-cross Abstract: Recently, Large Language Models (LLMs) have introduced a novel paradigm in Time Series Analysis (TSA), leveraging strong language capabilities

GAD in the Wild: Benchmarking Graph Anomaly Detection under Realistic Deployment Challenges

Model ReleasesDGX agent

arXiv:2605.07133v1 Announce Type: cross Abstract: Graph Anomaly Detection (GAD) is a critical task in graph machine learning with vital applications in financial fraud detection and social platform go

gamma-weakly heta-up-concavity: A Unified Framework for Non-Convex Optimization Beyond DR-Submodular and OSS Functions

ResearchDGX agent

arXiv:2602.13506v2 Announce Type: replace-cross Abstract: Optimizing non-convex functions is a fundamental challenge across machine learning and combinatorial optimization. We introduce and study gamm

GASim: A Graph-Accelerated Hybrid Framework for Social Simulation

SafetyDGX agent

arXiv:2605.07692v1 Announce Type: new Abstract: Large-scale social simulators are essential for studying complex social patterns. Prior work explores hybrid methods to scale up simulations, combining

Gated QKAN-FWP: Scalable Quantum-inspired Sequence Learning

Model ReleasesDGX agent

arXiv:2605.06734v1 Announce Type: cross Abstract: Fast Weight Programmers (FWPs) encode temporal dependencies through dynamically updated parameters rather than recurrent hidden states. Quantum FWPs (

GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2605.07817v1 Announce Type: cross Abstract: Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routin

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions

SafetyDGX agent

arXiv:2512.20974v3 Announce Type: replace-cross Abstract: Bayesian Reinforcement Learning (BRL), a subclass of Meta-Reinforcement Learning (Meta-RL), provides a principled framework for generalisation

Generalized Euler Logarithm and its Applications in Machine Learning: Natural Gradient, Backpropagation, Generalized EG, Mirror Descent and OLPS

Model ReleasesDGX agent

arXiv:2502.17500v3 Announce Type: replace-cross Abstract: This paper investigates in depth the fundamental properties of the two-parameter generalized Euler logarithm and its inverse, the associated d

Generative Modeling with Flux Matching

ResearchDGX agent

arXiv:2605.07319v1 Announce Type: cross Abstract: We introduce Flux Matching, a new paradigm for generative modeling that generalizes existing score-based models to a broader family of vector fields t

Geometric Kolmogorov--Arnold Network (GeoKAN)

Local AiDGX agent

arXiv:2605.06740v1 Announce Type: cross Abstract: We introduce Geometric Kolmogorov--Arnold Networks (GeoKANs), a family of geometry-aware KAN-type models in which approximation is carried out in lear

Globally Optimal Training of Spiking Neural Networks via Parameter Reconstruction

Model ReleasesDGX agent

arXiv:2605.08022v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) have been proposed as biologically plausible and energy-efficient alternatives to conventional Artificial Neural Networ

Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2410.06347v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) in robotics faces significant hurdles regarding sample efficiency and generalization across varying goals. While O

← Previous
1…279280281282283…358
Next →