AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
10 Jun 2026

A Comprehensive Survey of Direct Preference Optimization: Datasets, Theories, Variants, and Applications

SafetyDGX agent

arXiv:2410.15595v4 Announce Type: replace Abstract: With the rapid advancement of large language models (LLMs), aligning policy models with human preferences has become increasingly critical. Direct P

A Constrained Natural-Language Interface for Variational Multi-Physics Finite Element Simulations in FEniCS

Model ReleasesDGX agent

arXiv:2606.10928v1 Announce Type: cross Abstract: Large language models can reduce the manual effort required to set up finite element simulations, but they introduce reliability risks when generated

A Controlled Audit of Pretraining Contamination in Public Medical Vision-Language Benchmarks

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.10066v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) are evaluated on public benchmarks whose images and question-answer pairs have been freely downloadable for year

A History-Aware Visually Grounded Critic for Computer Use Agents

Model ReleasesDGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

A Note on the Strategic Confinement Problem

TutorialsDGX agent

arXiv:2606.09931v1 Announce Type: cross Abstract: Lampson's confinement problem asks how to prevent a program that processes confidential information from leaking it to a third party. We introduce the

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation

SafetyDGX agent

arXiv:2606.10366v1 Announce Type: cross Abstract: Simulation has become an essential tool for evaluating and improving vision-language-action (VLA) policies, offering scalable, reproducible, and contr

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

SafetyDGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

A Source Domain is All You Need: Source-Only Cross-OS Transfer Learning for APT Anomaly Detection via Semantic Alignment and Optimal Transport

SafetyDGX agent

arXiv:2606.10216v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) are stealthy, multi-stage cyberattacks whose detection is difficult due to scarce labeled traces, severe class imba

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

Model ReleasesDGX agent

arXiv:2505.01458v2 Announce Type: replace-cross Abstract: Navigation and manipulation are core capabilities in Embodied AI, but training agents to perform them directly in the real world is costly, ti

A Survey on Semantic Modeling for Building Energy Management

AgentsDGX agent

arXiv:2404.11716v2 Announce Type: replace Abstract: Building Energy Management (BEM) is central to reducing energy use and CO2 emissions in the building sector. Although IoT technologies now provide e

A Theory on Flow Matching with Neural Networks

ApplicationsDGX agent

arXiv:2606.10089v1 Announce Type: cross Abstract: In this work, we develop theoretical foundation for flow matching with neural-network-parameterized conditional velocity fields. We establish converge

A Unified Multi-Modal Framework for Intelligent Financial Systems: Integrating Reinforcement Learning, High-Frequency Trading, and Game-Theoretic Approaches with Cross-Modal Sentiment Analysis

SafetyDGX agent

arXiv:2606.10412v1 Announce Type: new Abstract: The rapid evolution of financial technology demands sophisticated artificial intelligence systems capable of handling diverse challenges across multiple

A Unified Siamese Learning Framework for Zero-Day Anomaly Detection and Classification in Optical Networks

ResearchDGX agent

arXiv:2606.10827v1 Announce Type: cross Abstract: A multi-similarity Siamese neural network unifies zero-day anomaly detection and one-shot classification in optical networks, achieving over 99% accur

A Unifying Lens on Supervised Fine-Tuning Through Target Distribution Design

TutorialsDGX agent

arXiv:2606.11189v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) typically maximizes the likelihood of every token in a demonstrated trajectory. However, an observed token can be non-uni

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

Model ReleasesDGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

Accelerating NeurASP with vectorization and caching

ResearchDGX agent

arXiv:2606.10787v1 Announce Type: new Abstract: Neurosymbolic AI combines neural networks with symbolic programs to create robust and explainable predictions. One such framework is NeurASP, which trai

Accounting for AI Inference in Corporate GHG Inventories: A Four-Tier Methodology for Scope 3 Category 1 Reporting

HardwareDGX agent

arXiv:2606.10660v1 Announce Type: cross Abstract: AI inference services -- API subscriptions, enterprise chat tools, and SaaS products with embedded AI features -- fall unambiguously within Scope 3 Ca

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design

Model ReleasesDGX agent

arXiv:2606.10493v1 Announce Type: cross Abstract: Local deployment of large Mixture-of-Experts (MoE) models falls short of the service quality achieved in cloud-scale environments, even under low-conc

ActiveMem: Distributed Active Memory for Long-Horizon LLM Reasoning

AgentsDGX agent

arXiv:2606.10532v1 Announce Type: new Abstract: Memory is essential for enabling large language model (LLM) agents to handle long-horizon reasoning tasks. Existing memory mechanisms are largely centra

Adoption of Generative Artificial Intelligence in the German Software Engineering Industry: An Empirical Study

SafetyDGX agent

arXiv:2601.16700v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) tools have seen rapid adoption among software developers. While adoption rates in the industry are

Advancing the State-of-the-Art in Empirical Privacy Auditing

Model ReleasesDGX agent

arXiv:2606.10481v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning of large language models (LLMs) can exhibit problematic memorization of individual training examples. Empirical privac

Aesthetic Perspectives in Information Systems Research: A Hermeneutic Analysis

TutorialsDGX agent

arXiv:2606.09839v1 Announce Type: cross Abstract: How might implicit aesthetic perspectives shape what Information Systems (IS) scholarship recognises as worthy of study (or not)? In this hermeneutic

Agentic Hybrid RAG for Evidence-Grounded Muon Collider Analysis

Model ReleasesDGX agent

arXiv:2606.10381v1 Announce Type: cross Abstract: Muon collider research spans accelerator physics, detector instrumentation, and high-energy phenomenology, with relevant evidence scattered across a r

Agentic Social Affordance Framework (ASAF): Agent Identity Design as a Collaboration Interface in Multi-Agent Systems

AgentsDGX agent

arXiv:2606.09832v1 Announce Type: cross Abstract: As AI systems evolve from single conversational agents to complex multi-agent architectures, a critical design dimension has been overlooked: how the

AI-Driven Analytics of Team-Teaching Talk: Acoustic Patterns across Experience, Cohorts and the Learning Design

ResearchDGX agent

arXiv:2606.09831v1 Announce Type: cross Abstract: As classroom cohorts expand, team teaching is increasingly used to integrate the expertise and pedagogical perspectives of multiple teachers. Yet, the

Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation

Model ReleasesDGX agent

arXiv:2606.09864v1 Announce Type: cross Abstract: Key-value (KV) cache quantization is widely used to reduce Large Language Model (LLM) inference memory, yet existing evaluations solely focus on measu

An Improved Generative Adversarial Network for Micro-Resistivity Imaging Logging Restoration

TutorialsDGX agent

arXiv:2606.10200v1 Announce Type: cross Abstract: An improved GAN-based imaging logging image restoration method is presented in this paper for solving the problem of partially missing micro-resistivi

An LLM-Native Psychometric Instrument Does Not Predict LLM Behavior: Evidence Across 25 Models

SafetyDGX agent

arXiv:2606.09843v1 Announce Type: cross Abstract: Large language models (LLMs) produce stable self-reports on personality inventories, but these self-reports do not predict observed behavior. Whether

Anomaly Detection and Root Cause Analysis for Microservice Systems

Model ReleasesDGX agent

arXiv:2606.09942v1 Announce Type: cross Abstract: Microservice systems are widely used to build cloud applications, yet their complexity makes failures inevitable, degrading user experience and causin

Architect-Ant: Editable Automatic Furnishing of Architectural Floor Plans

SafetyDGX agent

arXiv:2606.10953v1 Announce Type: new Abstract: Furnished floor plans are fundamental to real estate visualization, interior design, and architectural workflows. However, progress in automatic furnitu

ASA: Backbone-Training-Free Representation Engineering for Tool-Calling Agents

Model ReleasesDGX agent

arXiv:2602.04935v3 Announce Type: replace-cross Abstract: Adapting LLM agents to domain-specific tool calling remains notably brittle under evolving interfaces. Prompt and schema engineering is easy t

Assessing Automated Prompt Injection Attacks in Agentic Environments

Model ReleasesDGX agent

arXiv:2606.10525v1 Announce Type: cross Abstract: Indirect prompt injection poses a critical threat to LLM agents that interact with untrusted external data, yet automated attack methods--proven effec

Assessment of Personality Dimensions Across Situations in Dyadic Role-Play Scenarios

ResearchDGX agent

arXiv:2507.19137v2 Announce Type: replace-cross Abstract: Prior research indicates that users prefer assistive technologies whose personalities align with their own. This has sparked interest in autom

ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark

Model ReleasesDGX agent

arXiv:2505.23851v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to symbolic mathematics, yet existing evaluations often conflate pattern memorization wi

Atomic Intent Reasoning: Bringing LLM Semantics to Industrial Cross-Domain Recommendations

ApplicationsDGX agent

arXiv:2606.10357v1 Announce Type: cross Abstract: Cross-domain recommendation is a core problem in content-to-e-commerce platforms. Its objective is to leverage user interactions with content to infer

Attacks on Machine-Text Detectors Retain Stylistic Fingerprints

ResearchDGX agent

arXiv:2505.14608v3 Announce Type: replace-cross Abstract: Despite considerable progress in the development of machine-text detectors, the ease with which machine-text can be manipulated to evade detec

Attention-Discounted Adaptive Sampler for Masked Diffusion Language Models

ResearchDGX agent

arXiv:2606.10829v1 Announce Type: cross Abstract: Masked diffusion language models can reduce inference steps by revealing multiple tokens per denoising iteration, but this parallelism is fragile: pos

Attention Expansion: Enhancing Keyphrase Extraction from Long Documents with Attention-Augmented Contextualized Embeddings

Model ReleasesDGX agent

arXiv:2606.10716v1 Announce Type: cross Abstract: Pre-trained language models (PLMs) have achieved strong performance in keyphrase extraction (KPE), largely due to their ability to generate rich conte

au-Rec: A Verifiable Benchmark for Agentic Recommender Systems

Model ReleasesDGX agent

arXiv:2606.10156v1 Announce Type: cross Abstract: As recommender systems transition toward agentic, multi-turn conversational interfaces, evaluation paradigms have struggled to keep pace. Current benc

AuRA: Internalizing Audio Understanding into LLMs as LoRA

ResearchDGX agent

arXiv:2606.11033v1 Announce Type: cross Abstract: Recent efforts to extend large language models (LLMs) to speech inputs typically rely on cascaded ASR-LLM pipelines, end-to-end speech-language models

Automated Pronunciation Evaluation for Korean Toddler Speech using Speech Diarization and Self-Supervised Learning

ResearchDGX agent

arXiv:2606.10213v1 Announce Type: cross Abstract: Speech sound disorders affect approximately 44% of Korean pediatric communication disorder cases, yet automated assessment tools for Korean toddler sp

AutoPDE: Reliable Agentic PDE Solving via Explicitly Represented Solver Strategies

AgentsDGX agent

arXiv:2606.10752v1 Announce Type: new Abstract: Numerical solvers for partial differential equations (PDEs) are core computational tools in science and engineering. Building reliable PDE solvers requi

BadRobot: Jailbreaking Embodied LLM Agents in the Physical World

Model ReleasesDGX agent

arXiv:2407.20242v5 Announce Type: replace-cross Abstract: Embodied AI represents systems where AI is integrated into physical entities. Large Language Model (LLM), which exhibits powerful language und

Baseline-Free Policy Optimization for Neural Combinatorial Optimization

SafetyDGX agent

arXiv:2606.10321v1 Announce Type: cross Abstract: Neural combinatorial optimization (NCO) trains autoregressive policies to solve routing problems. The standard training algorithm, REINFORCE with a ro

Belief Acquisition as Stochastic Filtering

ResearchDGX agent

arXiv:2206.02178v3 Announce Type: replace Abstract: This paper studies how belief acquisition can be accomplished using stochastic filtering. First, a theoretical foundation for empirical beliefs is o

Belief-Space Control for Personalized Cancer Treatment via Active Inference

ResearchDGX agent

arXiv:2606.10376v1 Announce Type: new Abstract: Cancer treatment is at the core a sequential decision-making problem with partial observability, latent patient heterogeneity, and explicit constraints

Bellman-Taylor Score Decoding for Markov Decision Processes with State-Dependent Feasible Action Sets

SafetyDGX agent

arXiv:2606.10979v1 Announce Type: new Abstract: Many Markov decision processes (MDPs) in operations research have feasible actions that are state dependent and defined implicitly by various operationa

Benchmarking Knowledge Editing using Logical Rules

Model ReleasesDGX agent

arXiv:2606.10554v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in real-world applications that require access to up-to-date knowledge. However, retraining LLM

Between Amnesia and Chaos: A Memory Stability Expressivity Trilemma for Trainable Dissipative Oscillator Networks

ResearchDGX agent

arXiv:2606.09929v1 Announce Type: cross Abstract: Physical reservoir computing harnesses nonlinear mechanical dynamics but, by convention, freezes the substrate and trains only a linear readout, presu

Beyond Absolute Imitation: Anchored Residual Guidance for Privileged On-Policy Distillation

Local AiDGX agent

arXiv:2606.10385v1 Announce Type: cross Abstract: On-policy distillation (OPD) has demonstrated strong empirical gains in enhancing complex reasoning in LLMs by aligning a student model with a teacher

Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

Model ReleasesDGX agent

arXiv:2606.10803v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) excel at utilizing digital APIs and increasingly serve as the 'brain' of embodied AI, instructing robots to i

Beyond Static Evaluation: Co-Evolutionary Mechanisms for LLM-Driven Strategy Evolution in Adversarial Games

AgentsDGX agent

arXiv:2606.10389v1 Announce Type: new Abstract: Recent advances in LLM-driven code evolution have enabled automated discovery by iteratively generating and improving programs. However, applying these

Beyond Uniform Token-Level Trust Region in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.10968v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become standard for improving LLM reasoning. However, existing PPO-style trust-region mechan

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

Model ReleasesDGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression

ApplicationsDGX agent

arXiv:2606.10135v1 Announce Type: cross Abstract: Transitioning bidirectional video diffusion models into an autoregressive paradigm improves the interactivity of video world models, but existing caus

Blurry Window Attention

ResearchDGX agent

arXiv:2606.09862v1 Announce Type: cross Abstract: The Softmax Attention operation in Transformer language models has a quadratic complexity in the sequence length and a growing state size in the form

Boosting ECG Classification Performance by Pre-training with Synthesized Data

ApplicationsDGX agent

arXiv:2606.10802v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) typically require extensive datasets for effective training. In the medical domain, acquiring large-scale data is often ch

Building Change Detection in Earthquake: A Multi-Scale Interaction Network and A Change Detection Dataset

ResearchDGX agent

arXiv:2606.10329v1 Announce Type: cross Abstract: As one of the most destructive natural disasters, earthquakes have struck many countries around the world in recent years, causing serious economic lo

Business World Model

AgentsDGX agent

arXiv:2606.10044v1 Announce Type: new Abstract: Businesses are increasingly adopting AI-enabled tools to improve productivity, reduce costs, and enhance products and services. However, the transformat

Bypassing Copyright Protection in Diffusion-based Customization via Two-Stage Latent Feature Optimization

SafetyDGX agent

arXiv:2606.09909v1 Announce Type: cross Abstract: With the growing concerns over copyright infringement in diffusion-based customization, adversarial attacks have emerged as a prominent defense strate

← Previous
1…134135136137138…358
Next →