AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces

Model ReleasesDGX agent

arXiv:2605.08904v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and tool use. However, the fundamental cognitive faculties essential

Optimal FALQON for Quantum Approximate Optimization via Layer-wise Parameter Tuning

Model ReleasesDGX agent

arXiv:2605.08332v1 Announce Type: cross Abstract: Feedback-based adaptive quantum optimization (FALQON) is a promising approach for solving combinatorial problems on noisy intermediate-scale quantum (

Optimal Transport-Guided Adversarial Attacks on Graph Neural Network-Based Bot Detection

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.00318v2 Announce Type: replace-cross Abstract: The rise of bot accounts on social media poses significant risks to public discourse. To address this threat, modern bot detectors increasingl

Optimized Culprit Identification Using Mobilenet and Attention Mechanisms

Model ReleasesDGX agent

arXiv:2605.08169v1 Announce Type: cross Abstract: Automated culprit identification in surveillance systems is a critical task that requires high accuracy along with computational efficiency for real-t

Optimizer-Induced Mode Connectivity: From AdamW to Muon

ResearchDGX agent

arXiv:2605.09991v1 Announce Type: new Abstract: Mode connectivity has been widely studied, yet the role of the optimizer remains underexplored. We revisit it through optimizer-induced implicit regular

Oracle Poisoning: Corrupting Knowledge Graphs to Weaponise AI Agent Reasoning

Model ReleasesDGX agent

arXiv:2605.09822v1 Announce Type: cross Abstract: We define Oracle Poisoning, an attack class in which an adversary corrupts a structured knowledge graph that AI agents query at runtime via tool-use p

OracleTSC: Oracle-Informed Reward Hurdle and Uncertainty Regularization for Traffic Signal Control

Model ReleasesDGX agent

arXiv:2605.08516v1 Announce Type: new Abstract: Transparent decision-making is essential for traffic signal control (TSC) systems to earn public trust. However, traditional reinforcement learning-base

OrderFusion: Encoding Orderbook for End-to-End Probabilistic Intraday Electricity Price Forecasting

Model ReleasesDGX agent

arXiv:2502.06830v5 Announce Type: replace-cross Abstract: Probabilistic intraday electricity price forecasting is becoming increasingly important for short-term power-system operation. With increasing

Outlier-Robust Diffusion Solvers for Inverse Problems

ApplicationsDGX agent

arXiv:2605.09477v1 Announce Type: cross Abstract: Methods based on diffusion models (DMs) for solving inverse problems (IPs) have recently achieved remarkable performance. However, DM-based methods ty

Overcoming Multi-step Complexity in Multimodal Theory-of-Mind Reasoning: A Scalable Bayesian Planner

ResearchDGX agent

arXiv:2506.01301v2 Announce Type: replace Abstract: Theory-of-Mind (ToM) enables humans to infer mental states-such as beliefs, desires, and intentions-forming the foundation of social cognition. Howe

PAINET: A Principled Efficient Transformer for 3D Dynamics Modeling

ApplicationsDGX agent

arXiv:2510.04233v2 Announce Type: replace-cross Abstract: Modeling 3D dynamics is a fundamental problem in multi-body systems across scientific and engineering domains and has important practical impl

Pairwise is Not Enough: Hypergraph Neural Networks for Multi-Agent Pathfinding

Model ReleasesDGX agent

arXiv:2602.06733v2 Announce Type: replace-cross Abstract: Multi-Agent Path Finding (MAPF) is a representative multi-agent coordination problem, where multiple agents are required to navigate to their

PaperFit: Vision-in-the-Loop Typesetting Optimization for Scientific Documents

Model ReleasesDGX agent

arXiv:2605.10341v1 Announce Type: new Abstract: A LaTeX manuscript that compiles without error is not necessarily publication-ready. The resulting PDFs frequently suffer from misplaced floats, overflo

Parallel Multi-Circuit Quantum Feature Fusion in Hybrid Quantum-Classical Convolutional Neural Networks for Breast Tumor Classification

Model ReleasesDGX agent

arXiv:2512.02066v2 Announce Type: replace-cross Abstract: Quantum machine learning has emerged as a promising approach to improve feature extraction and classification tasks in high-dimensional data d

Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution

Model ReleasesDGX agent

arXiv:2605.09781v1 Announce Type: cross Abstract: Large Language Models exhibit mode collapse, producing homogeneous outputs that fail to explore valid solution spaces. We present QD-LLM, a framework

PARD-2: Target-Aligned Parallel Draft Model for Dual-Mode Speculative Decoding

ResearchDGX agent

arXiv:2605.08632v1 Announce Type: cross Abstract: Speculative decoding accelerates Large Language Models (LLMs) inference by using a lightweight draft model to propose candidate tokens that are verifi

parHSOM: A novel parallel Hierarchical Self-Organizing Map implementation

ResearchDGX agent

arXiv:2605.08164v1 Announce Type: cross Abstract: The digital age has completely transformed the way that information is processed and stored, which makes cybersecurity a crucial field of research. Cy

Path-Coupled Bellman Flows for Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.08253v1 Announce Type: cross Abstract: Distributional reinforcement learning (DRL) models the full return distribution, but existing finite-support or quantile-based methods rely on project

PathISE: Learning Informative Path Supervision for Knowledge Graph Question Answering

ResearchDGX agent

arXiv:2605.10791v1 Announce Type: new Abstract: Knowledge Graph Question Answering (KGQA) aims to answer user questions by reasoning over Knowledge Graphs (KGs). Recent KGQA methods mainly follow the

PATRA: Pattern-Aware Alignment and Balanced Reasoning for Time Series Question Answering

SafetyDGX agent

arXiv:2602.23161v2 Announce Type: replace Abstract: Time series reasoning demands both the perception of complex dynamics and logical depth. However, existing LLM-based approaches exhibit two limitati

PDEAgent-Bench: A Multi-Metric, Multi-Library Benchmark for PDE Solver Generation

Model ReleasesDGX agent

arXiv:2605.09636v1 Announce Type: new Abstract: PDE-to-solver code generation aims to automatically synthesize executable numerical solvers from partial differential equation (PDE) specifications. Thi

Perceptual Asymmetry Between Hue Categories: Evidence from Human Color Categorization

ResearchDGX agent

arXiv:2605.09339v1 Announce Type: cross Abstract: Human color categories are not uniformly distributed in perceptual space, yet most computational color models still assume fixed and evenly structured

Personalized Alignment Revisited: The Necessity and Sufficiency of User Diversity

Model ReleasesDGX agent

arXiv:2605.09119v1 Announce Type: cross Abstract: Personalized alignment aims to adapt large language models to heterogeneous user preferences, yet the precise theoretical conditions for its statistic

Personalizing LLMs with Binary Feedback: A Preference-Corrected Optimization Framework

SafetyDGX agent

arXiv:2605.10043v1 Announce Type: cross Abstract: Large Language Model (LLM) personalization aims to align model behaviors with individual user preferences. Existing methods often focus on isolated us

PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI

SafetyDGX agent

arXiv:2605.05682v2 Announce Type: replace-cross Abstract: Recent developments in AI safety research have called for red-teaming methods that effectively surface potential risks posed by generative AI

Phase Transitions in Affective Meaning Divergence: The Hidden Drift Before the Break

ResearchDGX agent

arXiv:2605.09043v1 Announce Type: cross Abstract: One partner says 'Fine' meaning resolution; the other hears surrender. The word is shared; the affective uptake is not. We formalize this as affective

PHMForge: Evaluating LLM Agents on Industrial Prognostics through MCP-Native, Algorithm-Grounded Tools

SafetyDGX agent

arXiv:2604.01532v2 Announce Type: replace Abstract: LLM agents are beginning to invoke industrial asset-management tools through the Model Context Protocol (MCP), yet whether they can act reliably on

Phoenix-VL 1.5 Medium Technical Report

Model ReleasesDGX agent

arXiv:2605.10391v1 Announce Type: cross Abstract: We introduce Phoenix-VL 1.5 Medium, a 123B-parameter natively multimodal and multilingual foundation model, adapted to regional languages and the Sing

PhyGround: Benchmarking Physical Reasoning in Generative World Models

Model ReleasesDGX agent

arXiv:2605.10806v1 Announce Type: cross Abstract: Generative world models are increasingly used for video generation, where learned simulators are expected to capture the physical rules that govern re

PhysHanDI: Physics-Based Reconstruction of Hand-Deformable Object Interactions

ApplicationsDGX agent

arXiv:2605.09538v1 Announce Type: cross Abstract: While existing methods for reconstructing hand-object interactions have made impressive progress, they either focus on rigid or part-wise rigid object

Physical probes expose and alleviate chemical-environment collapse in molecular representations

Local AiDGX agent

arXiv:2605.10429v1 Announce Type: cross Abstract: Nuclear magnetic resonance (NMR) spectroscopy provides an experimental readout of local chemical environments, but its use in molecular representation

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2605.09287v1 Announce Type: new Abstract: Large Language Model (LLM)-based search agents trained with reinforcement learning (RL) have significantly improved the performance of knowledge-intensi

PLACO: A Multi-Stage Framework for Cost-Effective Performance in Human-AI Teams

ResearchDGX agent

arXiv:2605.08388v1 Announce Type: new Abstract: Human-AI teams play a pivotal role in improving overall system performance when neither the human nor the model can achieve such performance on their ow

Playing games with knowledge: AI-Induced delusions need game theoretic interventions

Local AiDGX agent

arXiv:2605.08409v1 Announce Type: new Abstract: Conversational AI has a fundamental flaw as a knowledge interface: sycophantic chatbots induce epistemic entrenchment and delusional belief spirals even

Playing Games with My Heart: An Evaluation of AI Companion Apps

ApplicationsDGX agent

arXiv:2605.08093v1 Announce Type: cross Abstract: The use of chatbots for various forms of companionship is growing rapidly, raising a myriad of questions about simulated relationships, emotional depe

PnP-Corrector: A Universal Correction Framework for Coupled Spatiotemporal Forecasting

AgentsDGX agent

arXiv:2605.08935v1 Announce Type: new Abstract: Coupled spatiotemporal forecasting is important for predicting the future evolution of multiple interacting dynamical systems, such as in climate models

PoDAR: Power-Disentangled Audio Representation for Generative Modeling

ResearchDGX agent

arXiv:2605.10084v1 Announce Type: cross Abstract: The performance of audio latent diffusion models is primarily governed by generator expressivity and the modelability of the underlying latent space.

Policy Gradient Methods for Non-Markovian Reinforcement Learning

SafetyDGX agent

arXiv:2605.10816v1 Announce Type: cross Abstract: We study policy gradient methods for reinforcement learning in non-Markovian decision processes (NMDPs), where observations and rewards depend on the

Political Plasticity: An Analysis of Ideological Adaptability in Large Language Models

SafetyDGX agent

arXiv:2605.08415v1 Announce Type: new Abstract: Since the advent of Large Language Models (LLMs), a significant area of research has focused on their intrinsic biases, particularly in political discou

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2410.14702v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) exhibit impressive problem-solving abilities in various domains, but their visual comprehension and abstra

Portable Active Learning for Object Detection

TutorialsDGX agent

arXiv:2605.10349v1 Announce Type: cross Abstract: Annotating bounding boxes is costly and limits the scalability of object detection. This challenge is compounded by the need to preserve high accuracy

Position: Academic Conferences are Potentially Facing Denominator Gaming Caused by Fully Automated Scientific Agents

SafetyDGX agent

arXiv:2605.09915v1 Announce Type: cross Abstract: The implicit policy of maintaining relatively stable acceptance rates at top AI conferences, despite exponentially growing submissions, introduces a c

Position: AI Security Policy Should Target Systems, Not Models

Model ReleasesDGX agent

arXiv:2605.09504v1 Announce Type: cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, paral

Position: Avoid Overstretching LLMs for every Enterprise Task

ApplicationsDGX agent

arXiv:2605.09365v1 Announce Type: new Abstract: Enterprise workloads are dominated by deterministic, structured, and knowledge-dependent tasks operating under strict cost, latency, and reliability con

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead

Model ReleasesDGX agent

arXiv:2507.23009v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved remarkable results on a range of standardized tests originally designed to assess human cognitive a

Positional Encoding via Token-Aware Phase Attention

SafetyDGX agent

arXiv:2509.12635v3 Announce Type: replace-cross Abstract: We prove under practical assumptions that Rotary Positional Embedding (RoPE) introduces an intrinsic distance-dependent bias in attention scor

Positive Alignment: Artificial Intelligence for Human Flourishing

SafetyDGX agent

arXiv:2605.10310v1 Announce Type: new Abstract: Existing alignment research is dominated by concerns about safety and preventing harm: safeguards, controllability, and compliance. This paradigm of ali

PowerStep: Memory-Efficient Adaptive Optimization via ell_p-Norm Steepest Descent

ResearchDGX agent

arXiv:2605.10335v1 Announce Type: cross Abstract: Adaptive optimizers, most notably Adam, have become the default standard for training large-scale neural networks such as Transformers. These methods

PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models

Model ReleasesDGX agent

arXiv:2605.08800v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) may memorize sensitive cross-modal information during pretraining. However, existing MLLM unlearning benchmar

Practical Wi-Fi-based Motion Recognition Under Variable Traffic Patterns

ResearchDGX agent

arXiv:2605.08308v1 Announce Type: cross Abstract: Wi-Fi sensing detects human motions and activities by analysing the channel state information (CSI) derived from Wi-Fi transmissions. However, the imp

PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2602.03190v3 Announce Type: replace-cross Abstract: Reinforcement learning algorithms such as group-relative policy optimization (GRPO) have shown strong potential for improving the mathematical

Prediction Bottlenecks Don't Discover Causal Structure (But Here's What They Actually Do)

Model ReleasesDGX agent

arXiv:2605.09169v1 Announce Type: cross Abstract: A Mamba state-space model trained only for next-step prediction appears to recover Granger-causal structure through a simple readout S = |W_{out} W_{i

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Model ReleasesDGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

Pretraining large language models with MXFP4

Model ReleasesDGX agent

arXiv:2605.09825v1 Announce Type: cross Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We a

Preventing Rank Collapse in Federated Low-Rank Adaptation with Client Heterogeneity

Local AiDGX agent

arXiv:2602.13486v2 Announce Type: replace-cross Abstract: Federated low-rank adaptation (FedLoRA) has facilitated communication-efficient and privacy-preserving fine-tuning of foundation models for do

Primal-Dual Guided Decoding for Constrained Discrete Diffusion

SafetyDGX agent

arXiv:2605.09749v1 Announce Type: new Abstract: Discrete diffusion models generate structured sequences by progressively unmasking tokens, but enforcing global property constraints during generation r

PrimeKG-CL: A Continual Graph Learning Benchmark on Evolving Biomedical Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.10529v1 Announce Type: new Abstract: Biomedical knowledge graphs underwrite drug repurposing and clinical decision support, yet the upstream ontologies they depend on update on independent

Priming: Hybrid State Space Models From Pre-trained Transformers

Model ReleasesDGX agent

arXiv:2605.08301v1 Announce Type: cross Abstract: Hybrid State-Space models combine Attention with recurrent State-Space Model (SSM) layers, balancing eidetic memory from Attention with compressed fad

PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines

Model ReleasesDGX agent

arXiv:2605.10614v1 Announce Type: new Abstract: Multi-agent LLM systems introduce a security risk in which sensitive information accessed by one agent can propagate through shared context and reappear

Privacy-Aware Video Anomaly Detection through Orthogonal Subspace Projection

SafetyDGX agent

arXiv:2605.08651v1 Announce Type: cross Abstract: Video anomaly detection (VAD) systems often prioritize accuracy while overlooking privacy concerns, limiting their suitability for real-world deployme

← Previous
1…271272273274275…358
Next →