AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
8 Jul 2026

Learnable Weighting of Intra-Attribute Distances for Categorical Data Clustering with Nominal and Ordinal Attributes

ResearchDGX agent

arXiv:2607.05464v1 Announce Type: cross Abstract: The success of categorical data clustering generally much relies on the distance metric that measures the dissimilarity degree between two objects. Ho

Learning 4D Geometric Priors for Inference-Efficient World Action Models

ApplicationsDGX agent

arXiv:2607.05468v1 Announce Type: cross Abstract: World Action Models (WAMs) have shown strong potential for robotic manipulation by jointly modeling visual future dynamics and executable action seque

Learning The Minimum Action Distance

AgentsDGX agent

arXiv:2506.09276v4 Announce Type: replace-cross Abstract: This paper presents a state representation framework for Markov decision processes (MDPs) that can be learned solely from state trajectories,


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.05458v1 Announce Type: cross Abstract: Large language model (LLM) agents are usually improved by changing prompts, models, or hand-written workflows, while the execution harness around the

LEGATO 2: Toward Multimodal Sheet Music Recognition and Understanding

ResearchDGX agent

arXiv:2607.05769v1 Announce Type: cross Abstract: We propose a novel pipeline, Legato 2, for extracting symbolic notation and semantic knowledge from images of sheet music. Legato 2 features the first

Lingering Authority: Revocable Resource-and-Effect Capabilities for Coding Agents

SafetyDGX agent

arXiv:2606.22504v1 Announce Type: cross Abstract: Coding agents often receive broad tool access for an entire task, even when a resource is needed only for one subgoal. We call this gap lingering auth

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability

Model ReleasesDGX agent

arXiv:2607.06157v1 Announce Type: cross Abstract: Deliberation plays a crucial role in collaboration; when humans work together, they naturally engage in communication to align information and reach a

LLM-Guided Measurement Credibility Correction for Trustworthy Industrial Process Inference

Local AiDGX agent

arXiv:2607.06111v1 Announce Type: cross Abstract: Industrial prediction and soft sensing depend on credible input measurements. In field deployment, a predictor may receive biased, delayed, stale, or

LongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction Synthesis

Model ReleasesDGX agent

arXiv:2607.06160v1 Announce Type: cross Abstract: Synthesizing long-context supervised fine-tuning (SFT) data is a scalable way to enhance the long-context understanding of large language models (LLMs

MCP-Enabled Agentic AI for Autonomous IPoDWDM Network Lifecycle Automation

AgentsDGX agent

arXiv:2607.05975v1 Announce Type: cross Abstract: This demo presents an MCP-enabled agentic AI architecture for autonomous control of vendor-agnostic IPoDWDM networks. We demonstrate live end-to-end l

Medix: Out-of-Distribution Detection from Unlabeled Wild Data via Robust Gradient Statistics

ApplicationsDGX agent

arXiv:2510.06505v2 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection plays a crucial role in ensuring the robustness of machine learning systems deployed in real-world applica

Memory in the Loop: In-Process Retrieval as ExtendedWorking Memory for Language Agents

Model ReleasesDGX agent

arXiv:2607.05690v1 Announce Type: new Abstract: Language agents run a loop - observe, reason, act - but the memory they reason over sits outside it: a store queried at most once per turn. We study the

MLLM-LLaVA-FL: Multimodal Large Language Model Assisted Federated Learning

Local AiDGX agent

arXiv:2409.06067v3 Announce Type: replace Abstract: Previous studies on federated learning (FL) often encounter performance degradation due to data heterogeneity among different clients. In light of t

Modality Relevance is not Modality Utility: Post-hoc Selective Modality Escalation for Cost-Aware Multimodal RAG

ResearchDGX agent

arXiv:2607.05438v1 Announce Type: cross Abstract: Multimodal retrieval-augmented generation (RAG) grounds a generator in evidence drawn from heterogeneous modalities -- text, tables, and images. The d

Most LLM Conformity Needs No Speaker: Measuring the Speaker-Free Floor in Peer-Pressure Benchmarks

ResearchDGX agent

arXiv:2607.05545v1 Announce Type: cross Abstract: LLM conformity is often used to describe cases where a model changes a correct answer toward a peer or group response. We show that most of this appar

Multi-Agent Deep Reinforcement Learning for Multi Objective Battery Management in Dairy Farms

AgentsDGX agent

arXiv:2607.06489v1 Announce Type: new Abstract: The dairy industry in Ireland has a large potential for the integration of renewable energy and the reduction of carbon emissions. However, researchers

Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic SpeechLLMs

Model ReleasesDGX agent

arXiv:2601.12494v3 Announce Type: replace-cross Abstract: Audio large language models (LLMs) enable unified speech understanding and generation, but adapting them to linguistically complex and dialect

Narrative-Centered Emotional Reflection: An Early Prototype for AI-Supported Emotional Self-Reflection

AgentsDGX agent

arXiv:2504.20342v2 Announce Type: replace-cross Abstract: Reflexion is an AI-powered prototype designed to explore structured emotional self-reflection. By integrating emotion detection, layered refle

Narrative World Model: Narratology-Grounded Writer Memory for Long-Form Fiction

Model ReleasesDGX agent

arXiv:2607.05577v1 Announce Type: new Abstract: Long-form fiction writers need memory that answers multi-hop questions about evolving story state: who knows a secret and when they learned it, whether

NegROI: Click-Centric Uncertainty-Guided Refinement with Scene-Conditioned Negative Prompts for Robust Interactive 3D Segmentation

Model ReleasesDGX agent

arXiv:2607.05955v1 Announce Type: cross Abstract: Interactive 3D segmentation aims to extract object masks in point clouds with minimal user clicks. Despite recent progress, most existing approaches s

Onnes: A Physics-Grounded Multi-Agent LLM Simulator for Cryogenic Fault Diagnosis in Quantum Computing Infrastructure

Model ReleasesDGX agent

arXiv:2607.05805v1 Announce Type: new Abstract: Dilution refrigerators are the enabling infrastructure of superconducting quantum computers, yet their fault diagnosis is still dominated by threshold a

OpenGround: Planning-based Online Perception for Open-World 3D Visual Grounding

ResearchDGX agent

arXiv:2512.23020v3 Announce Type: replace-cross Abstract: 3D visual grounding aims to locate objects based on natural language descriptions in 3D scenes. Existing supervised methods are limited by gen

PatchOptic for Shared-State LLM Workflows with Projected Views and Verified Structured Updates

Model ReleasesDGX agent

arXiv:2607.05483v1 Announce Type: cross Abstract: Agentic workflows often operate over shared, structured state. Because LLM context windows are limited, each model invocation is typically shown only

PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation

Model ReleasesDGX agent

arXiv:2607.05915v1 Announce Type: new Abstract: PCB routing is the task of connecting the nets of a board with copper traces under strict design rules, yet learning-based methods still lag behind rule

Perceptually Aligning Representations of Music via Noise-Augmented Autoencoders

ResearchDGX agent

arXiv:2511.05350v3 Announce Type: replace-cross Abstract: We argue that training autoencoders to reconstruct inputs from noised versions of their encodings, when combined with perceptually motivated l

Physics-Regularized Machine Learning for Proprioceptive Vehicle Localization Using Onboard Sensors

Local AiDGX agent

arXiv:2607.05663v1 Announce Type: cross Abstract: Accurate and robust localization is essential for autonomous mobility systems in real-world environments. While fusing Inertial Measurement Unit (IMU)

Pitwall: Faithful Natural-Language Race-Strategy Briefings from a Calibrated Real-Time Monte Carlo Engine

ApplicationsDGX agent

arXiv:2607.06495v1 Announce Type: cross Abstract: Live sports commentary is grounded generation under a deadline: statements concern real, named athletes, the grounding state changes every few seconds

Plainbook: Data Science, in Plain Language

ResearchDGX agent

arXiv:2607.05717v1 Announce Type: cross Abstract: Jupyter Notebooks have become widely adopted in data science, as they allow the sharing of reproducible computational analysis. They are, however, acc

Platonic Representations for Poverty Mapping: Unified Vision-Language Codes or Agent-Induced Novelty?

SafetyDGX agent

arXiv:2508.01109v3 Announce Type: replace Abstract: We investigate whether socioeconomic indicators, like household wealth, leave recoverable informational imprints in both satellite imagery (capturin

PluraMath: Extending Mathematical Reasoning Evaluation Beyond High-Resource Languages

Model ReleasesDGX agent

arXiv:2607.05992v1 Announce Type: cross Abstract: Mathematical reasoning has become a central task for evaluating and tuning reasoning Large Language Models (LLMs), yet existing benchmarks remain heav

PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails

Model ReleasesDGX agent

arXiv:2607.05910v1 Announce Type: cross Abstract: Image guardrails are typically trained and evaluated under a fixed safety policy, implicitly treating safety as an intrinsic property of an image. Rea

PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2607.06008v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external env

PORTS: Preference-Optimized Retrievers for Tool Selection with Large Language Models

SafetyDGX agent

arXiv:2607.05441v1 Announce Type: cross Abstract: Integrating external tools with Large Language Models (LLMs) has emerged as a promising paradigm for accomplishing complex tasks. Since LLMs still str

Position: EU AI Act's Research Exemptions Can Break the Publication Norms of Major AI Conferences

SafetyDGX agent

arXiv:2506.03218v2 Announce Type: replace-cross Abstract: The EU has become one of the vanguards in regulating the digital age. A particularly important regulation in the Artificial Intelligence (AI)

Position: Preventing AI-Generated CSAM Necessitates New Approaches to AI Safety

SafetyDGX agent

arXiv:2607.05407v1 Announce Type: cross Abstract: Modern artificial intelligence (AI) systems present profound new risks to child safety. AI is increasingly being misused to create AI-generated child

Privilege and confidentiality in generative AI workflows

Model ReleasesDGX agent

arXiv:2607.05479v1 Announce Type: cross Abstract: Generative AI (GenAI) systems store and process client data in three distinct ways: in the model's parameters through training and memorisation, in th

Prompt-Adapter Context Routing for Parameter-Efficient Multi-Shot Long Video Extrapolation

Model ReleasesDGX agent

arXiv:2607.06481v1 Announce Type: cross Abstract: We present PACR-Video, a parameter-efficient framework for multi-shot long video extrapolation that preserves recurring entities, scene structure, vis

Prompt Coach: An Empirical Evaluation of an Agentic Tutor for Learning Prompt Engineering in Software Development

AgentsDGX agent

arXiv:2607.06074v1 Announce Type: cross Abstract: Prompt engineering has emerged as a critical yet undertaught skill for software developers, one that traditional learning approaches are ill-equipped

Prompt Robustness Is Task-Dependent: Comparing Objective and Belief-Style Questions in LLM Evaluation

ResearchDGX agent

arXiv:2607.05554v1 Announce Type: cross Abstract: Survey-style evaluations of large language models often treat a prompted response as a measure of a model's values or beliefs. This assumption is part

Prompt-to-Paper: Agentic AI System for Bioinformatics

AgentsDGX agent

arXiv:2607.05456v1 Announce Type: new Abstract: While recent advances in large language models have enabled end-to-end automated manuscript generation, existing systems suffer from three critical defi

Proof of Execution: Runtime Verification for Governed AI Agent Actions

AgentsDGX agent

arXiv:2607.05397v1 Announce Type: cross Abstract: Agent systems increasingly execute rather than advise. When an AI agent queries regulated data, invokes effectful tools, and mutates persistent state,

Property-Driven Synthetic Data Engineering for Data-Scarce Software Systems: Reflections from the Breast Cancer Domain

SafetyDGX agent

arXiv:2607.06133v1 Announce Type: cross Abstract: Modern software systems increasingly depend on data for analysis, prediction, testing, and decision-making. Yet many important domains, including medi

Propose and Attend: Training-free MLLM Grounding Confidence via Multi-Token Localized Attention

Local AiDGX agent

arXiv:2607.05978v1 Announce Type: cross Abstract: Multimodal large language models can emit localized predictions, bounding boxes for objects and temporal windows for video and audio events, but they

Provable learning separation for predicting time-evolution of quantum many-body systems

ResearchDGX agent

arXiv:2607.06472v1 Announce Type: cross Abstract: Given that quantum computers are naturally suited to simulate the behavior of quantum many-body systems, an immediate question arises: can one formula

PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet

ResearchDGX agent

arXiv:2607.06097v1 Announce Type: cross Abstract: 3D dense captioning, an emerging vision-language task, aims to generate descriptive sentences for each object in the 3D scene. Despite the impressive

Quantifying Frontier LLM Capabilities for Container Sandbox Escape

Model ReleasesDGX agent

arXiv:2603.02277v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly act as autonomous agents, using tools to execute code, read and write files, and access networks, cr

Rendering-Aware Bayesian 3D Gaussian Splatting with Native Uncertainty and Adaptive Complexity Control

ResearchDGX agent

arXiv:2607.05522v1 Announce Type: cross Abstract: 3D Gaussian splatting (3DGS) is a strong representation for real-time novel-view synthesis, but its standard training pipeline relies on point estimat

Replication in Visual Diffusion Models: A Survey and Outlook

ApplicationsDGX agent

arXiv:2408.00001v2 Announce Type: replace-cross Abstract: Visual diffusion models have revolutionized the field of creative AI, producing high-quality and diverse content. However, they inevitably mem

ResonatorLM: Causal Resonant Field Mixing for Efficient Long-Context Language Modelin

ResearchDGX agent

arXiv:2607.05583v1 Announce Type: cross Abstract: Contemporary language models are dominated by the transformer architecture, which leverages self-attention mechanisms to enable more efficient, parall

Responsible Personalisation: The Double-Edged Sword of Personalisation in Human-Robot Interaction

ResearchDGX agent

arXiv:2607.06344v1 Announce Type: cross Abstract: While personalisation is becoming a defining capability in human-robot interaction (HRI), the existing literature on responsible personalisation remai

Rethinking Indic AI from a Lens of Cultural Heritage Preservation

TutorialsDGX agent

arXiv:2607.06544v1 Announce Type: new Abstract: As Artificial Intelligence (AI) makes inroads into different parts of the Indian subcontinent, there is significant interest in studying how AI impacts

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance

ResearchDGX agent

arXiv:2509.23876v3 Announce Type: replace-cross Abstract: Autoregressive (AR) models based on next-scale prediction have emerged as a powerful tool for image generation, but they face a critical weakn

Reward as An Agent for Embodied World Models

AgentsDGX agent

arXiv:2606.19990v2 Announce Type: replace Abstract: While RL has become a promising tool for refining world models, existing methods largely rely on conservative rollouts near the training distributio

Reward-Density Heuristic for Dynamic Multi-Vehicle Routing: Performance and Computational Efficiency

AgentsDGX agent

arXiv:2607.06066v1 Announce Type: new Abstract: The Vehicle Routing Problem (VRP) and its variants represent some of the most practically consequential optimization challenges in modern logistics and

RMISC: A Large-scale Real-world Multivariate Corpus for Time Series Foundation Models

ApplicationsDGX agent

arXiv:2607.06504v1 Announce Type: new Abstract: Recent years have witnessed the emergence of multivariate modeling using time series foundation models (TSFMs), which achieve advanced zero-shot general

RoME: Robust Mixture of Low-Rank Experts against Multiple Adversarial Perturbations

TutorialsDGX agent

arXiv:2607.06109v1 Announce Type: cross Abstract: Multi-perturbation adversarial training (MAT) aims to achieve robustness against multiple ell_p perturbations but suffers from robustness trade-offs b

RPAM: A Principled Metric for Evaluating Associations in Language Models with High Predictive Validity in Downstream Outputs

Model ReleasesDGX agent

arXiv:2607.05679v1 Announce Type: cross Abstract: Language models (LMs) exhibit problematic biases, such as stereotypes. Effectively analyzing and mitigating such biases requires accurate and generali

RSF-GLLM: Bridging the Semantic Gap in Multi-Hop Knowledge Graph QA via Recurrent Soft-Flow and Decoupled LLM Generation

ResearchDGX agent

arXiv:2607.06527v1 Announce Type: cross Abstract: Multi-hop Question Answering over Knowledge Graphs faces a critical challenge: traditional retrieve-then-read pipelines break differentiability, preve

RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications

Model ReleasesDGX agent

arXiv:2607.06411v1 Announce Type: cross Abstract: Developers increasingly delegate real maintenance work to product-grade coding agents, and many state tasks in their native language, in the style of

Safe Bayesian Optimization with Counterfactual Policies

SafetyDGX agent

arXiv:2607.05620v1 Announce Type: cross Abstract: In many decision-making settings, new interventions are acceptable only if they do not reduce outcomes below some established threshold. For example,

← Previous
1…8182838485…358
Next →