AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
16 Jul 2026

Efficient and Privacy Aware Edge Cloud Collaborative Inference for Large Language Models

Local AiDGX agent

arXiv:2607.13093v1 Announce Type: cross Abstract: On-device LLM inference faces a trilemma of response latency, limited hardware resources and user privacy. Full cloud inference delivers strong comput

Efficient Text-to-Audio Generation via Pruning

Model ReleasesDGX agent

arXiv:2607.13330v1 Announce Type: cross Abstract: Diffusion-based text-to-audio generative models such as AudioLDM achieve high perceptual quality and strong semantic consistency; however, their pract

EMAGN: Efficient Multi-Attention Graph Network via Learned Clustering for Scalable Traffic Forecasting

HardwareDGX agent

arXiv:2607.13241v1 Announce Type: cross Abstract: Traffic forecasting is highly challenging due to complex and nonlinear spatial and temporal dependencies. Self-attention mechanisms have been widely a


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-Loop Table Recognition

SafetyDGX agent

arXiv:2607.13347v1 Announce Type: cross Abstract: LLM-as-a-judge is widely used to provide feedback and selection signals in closedloop regeneration, but this use remains insufficiently validated. We

Experience Memory Graph: One-Shot Error Correction for Agents

AgentsDGX agent

arXiv:2607.13884v1 Announce Type: new Abstract: Large Language Model (LLM) agents have shown remarkable capabilities in autonomous decision-making by generating sequential trajectories of states, acti

Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Internal Audit Perspective

ResearchDGX agent

arXiv:2607.13469v1 Announce Type: cross Abstract: The banking sector increasingly relies on automated systems to monitor electronic transactions for signs of fraud, yet conventional rule-based approac

Explaining Reinforcement Learning Agents via Inductive Logic Programming

SafetyDGX agent

arXiv:2607.13655v1 Announce Type: new Abstract: Explainable Reinforcement Learning (XRL) seeks to make Reinforcement Learning (RL) policies more transparent and interpretable, a key requirement in saf

ExTernD: Expanded-Rank Ternary Decomposition Ternary LLM PTQ with Accuracy Approaching Any Quantization Level

Model ReleasesDGX agent

arXiv:2607.13511v1 Announce Type: cross Abstract: We introduce ExTernD (Expanded-rank Ternary Decomposition), a post-training factorization of each LLM weight matrix A in R^{m imes n} into A approx B

EZSMT Version 3, Matured

ResearchDGX agent

arXiv:2607.13344v1 Announce Type: new Abstract: Constraint Answer Set Programming (CASP) is a hybrid reasoning paradigm that combines Answer Set Programming (ASP) with Constraint Processing and Satisf

Faithful Autoformalization of Natural Language Assertions

ResearchDGX agent

arXiv:2607.13303v1 Announce Type: cross Abstract: Formal contracts are essential for software testing and verification, yet writing them remains labor-intensive and error-prone. LLMs offer a promising

Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open Challenges

Local AiDGX agent

arXiv:2607.13045v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as a key paradigm for privacy-preserving collaborative model training across distributed and heterogeneous data so

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

SafetyDGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

FixItFlow: Automated Troubleshooting Guide Generation from Cloud Incidents

TutorialsDGX agent

arXiv:2607.13035v1 Announce Type: cross Abstract: Cloud services experience frequent incidents that require rapid diagnosis and resolution. Troubleshooting guides help engineers respond consistently,

From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of Mobile Robots Using RGB-D Perception

Model ReleasesDGX agent

arXiv:2607.13624v1 Announce Type: cross Abstract: Natural language interaction provides an intuitive way for non-expert users to communicate with robotic platforms. However, transforming user requests

From Prediction to Collaboration: Interactive Symbolic Music Analysis

Model ReleasesDGX agent

arXiv:2607.13587v1 Announce Type: cross Abstract: Automatic symbolic music analysis has made substantial progress, yet existing systems are typically designed for a single mode of use, such as full-sc

Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit

HardwareDGX agent

arXiv:2607.13095v1 Announce Type: cross Abstract: We present a full-pipeline inference optimization for the MiMo-V2.5 model family, which combines Hybrid Sliding Window Attention (Hybrid SWA), sparse

Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code

TutorialsDGX agent

arXiv:2607.13921v1 Announce Type: cross Abstract: Languages with rich static semantics, such as Rust, provide stronger guarantees for AI-generated code, but their strictness makes generation more diff

GeoAnchor: Collaborative Reasoning via Latent Decomposition for 3D Spatial Understanding

Local AiDGX agent

arXiv:2607.13454v1 Announce Type: cross Abstract: Although multimodal large language models (MLLMs) have achieved remarkable progress, understanding 3D spatial relationships from 2D images remains a c

GHR-VLM: Making Zero-Shot Transit Video Analytics Realizable with Grounded Hybrid Reasoning

ApplicationsDGX agent

arXiv:2607.13569v1 Announce Type: cross Abstract: Transit video understanding can provide valuable fine-grained data that conventional passenger counters and fare systems cannot capture. However, supe

Greedy Volume Maximization of Gradient Embeddings for Long-Tailed Frame-Level Bioacoustic Active Learning

ResearchDGX agent

arXiv:2607.13555v1 Announce Type: cross Abstract: Bioacoustic call-type classification relies on costly expert annotation. Active learning can reduce this burden by selecting a small batch of segments

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs

SafetyDGX agent

arXiv:2607.13712v1 Announce Type: cross Abstract: Despite the rapid progress of Multimodal Large Language Models (MLLMs), they still suffer from untruthfulness issues, such as visual hallucinations, c

Grounded world models in biological organisms and future embodied AI

AgentsDGX agent

arXiv:2607.13560v1 Announce Type: cross Abstract: Recent advances in generative and embodied AI have been driven by large-scale predictive learning over multimodal data. However, the resulting systems

Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable

Local AiDGX agent

arXiv:2607.13285v1 Announce Type: new Abstract: The capability of a modern AI agent depends not only on its foundation model but also on its harness, which constructs prompts, manages state, invokes t

HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation

SafetyDGX agent

arXiv:2607.09776v2 Announce Type: replace-cross Abstract: When adapting Vision Language Action (VLA) models to downstream tasks, multiple rounds of post-training are often required to progressively ad

How Agents Ask for Permission: User Permissions for AI Agents, from Interfaces to Enforcement

AgentsDGX agent

arXiv:2607.13718v1 Announce Type: cross Abstract: As AI agents gain prevalance, users are increasingly exposed to the risks such systems entail. Prompt injection attacks, as well as hallucination, can

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

Model ReleasesDGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

HRO: Hierarchical Room-to-Object Framework for Zero-Shot Object Goal Navigation with Large Language Models

Local AiDGX agent

arXiv:2607.13072v1 Announce Type: cross Abstract: Zero-shot object-goal navigation aims to enable an intelligent agent to explore and navigate to objects of unknown categories in an unfamiliar environ

Human4K: A Large-Scale 4K Multi-View Mocap Dataset for Whole-Body 3D Human Reconstruction

SafetyDGX agent

arXiv:2607.13646v1 Announce Type: cross Abstract: Recent advances in 3D human reconstruction have improved overall performance, yet current models still fail in the most challenging real-world scenari

IMMNet: Hybrid Fusion of Model-based and Data-driven Approaches for Maneuvering Target Tracking

TutorialsDGX agent

arXiv:2607.13573v1 Announce Type: cross Abstract: Maneuvering target tracking in three-dimensional space remains a challenging problem due to complex motion dynamics and model mismatch. To address thi

Improving Molecular Property Prediction in Small Language Models Using Graph-based Tools

AgentsDGX agent

arXiv:2607.13115v1 Announce Type: new Abstract: Small language models (SLMs) have shown promise for zero-shot molecular property prediction from SMILES strings, yet they often suffer from structural b

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models

Model ReleasesDGX agent

arXiv:2607.13408v1 Announce Type: cross Abstract: Recent text-to-audio models generate high-quality audio, but often fail to follow instructions involving multiple sound events and temporal order. Thi

Improving Wind and Solar Power Prediction with Efficient Wrapper-based Feature Selection: An Empirical Study

ApplicationsDGX agent

arXiv:2607.14024v1 Announce Type: cross Abstract: With rising global energy demand and growing awareness of climate change and its impacts, the share of renewable energies in the global energy mix con

Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs

Model ReleasesDGX agent

arXiv:2607.13080v1 Announce Type: cross Abstract: Autonomous coding agents force engineering organizations to choose between API-based frontier models -- strong reasoning at high token cost -- and on-

Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate

Model ReleasesDGX agent

arXiv:2510.10002v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in sensitive everyday contexts -- offering personal advice, mental health support, and mor

Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thought via Predicate Substitution

Model ReleasesDGX agent

arXiv:2607.13069v1 Announce Type: new Abstract: Large language models produce chain-of-thought (CoT) reasoning that appears logically sound yet may not genuinely depend on its stated premises. We intr

Introducing Human-Centeredness in AI-Assisted Lexicography

SafetyDGX agent

arXiv:2607.11808v2 Announce Type: replace-cross Abstract: This paper proposes a human-centered artificial intelligence (HCAI) framework for AI-assisted lexicography. While generative AI offers signifi

Inverse-LLaVA: Rethinking Multimodal Alignment via Text-to-Vision Mapping

SafetyDGX agent

arXiv:2508.12466v2 Announce Type: replace-cross Abstract: Traditional multimodal learning approaches rely on alignment pre-training to bridge vision and language modalities, typically by projecting vi

Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for Structured Data Classification

Model ReleasesDGX agent

arXiv:2607.13413v1 Announce Type: cross Abstract: This study presents an empirical benchmarking comparison between Kolmogorov-Arnold Networks (KANs) and Multi-Layer Perceptrons (MLPs) on structured ta

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations

ResearchDGX agent

arXiv:2607.13770v1 Announce Type: cross Abstract: Video diffusion transformers (vDiTs) generate high quality video but introduce extremely high compute cost due to the long diffusion timesteps and sel

Koopman-driven grip force prediction through EMG sensing

ResearchDGX agent

arXiv:2409.17340v2 Announce Type: replace-cross Abstract: Loss of hand function due to conditions like stroke or multiple sclerosis significantly impacts daily activities. Robotic rehabilitation provi

LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search Reasoning

SafetyDGX agent

arXiv:2607.13501v1 Announce Type: new Abstract: Reinforcement learning for multi-turn search reasoning typically relies on terminal outcome rewards, which cannot distinguish useful, redundant, and har

Learning Engagement Assistant (LEA): Cross-Course Scalability and Classroom Evaluation of an Agentic AI Tutoring System

AgentsDGX agent

arXiv:2607.13370v1 Announce Type: cross Abstract: This paper is an extension of a paper presented at the ICAART 2026 conference, which introduced LEA (Learning Engagement Assistant), an adaptive AI tu

Learning Physics-Guided Residual Dynamics for Deformable Object Simulation

ApplicationsDGX agent

arXiv:2607.13451v1 Announce Type: cross Abstract: Simulating deformable objects is essential for a wide range of robotic manipulation applications, yet accurately predicting their dynamics remains cha

Learning Red Agent Policy from Observations for Neurosymbolic Autonomous Cyber Agents

SafetyDGX agent

arXiv:2606.18223v2 Announce Type: replace-cross Abstract: With sophisticated cyber-attacks becoming increasingly prevalent, modern networks require intelligent autonomous cyber-defense agents trained

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

SafetyDGX agent

arXiv:2607.13172v1 Announce Type: new Abstract: We address the problem of safely training an agent policy and deploying a good and safe policy, in settings where the environment dynamics are unknown a

Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies

SafetyDGX agent

arXiv:2604.00830v3 Announce Type: replace-cross Abstract: Test-Time Learning (TTL) enables language agents to iteratively refine their performance through repeated interactions with the environment at

Left-right asymmetry in predicting brain activity from LLMs' representations emerges with their formal linguistic competence

ResearchDGX agent

arXiv:2602.12811v2 Announce Type: replace-cross Abstract: When humans and large language models (LLMs) process the same text, activations in the LLMs correlate with brain activity measured, e.g., with

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

Model ReleasesDGX agent

arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to s

LLM-Guided Reinforcement Learning for Audio-Visual Speech Enhancement

ResearchDGX agent

arXiv:2603.13952v3 Announce Type: replace-cross Abstract: In existing Audio-Visual Speech Enhancement (AVSE) methods, objectives such as Scale-Invariant Signal-to-Noise Ratio (SI-SNR) and Mean Squared

MASPRM: Multi-Agent System Process Reward Model

SafetyDGX agent

arXiv:2510.24803v3 Announce Type: replace-cross Abstract: Inference-time search over multi-agent systems (MAS) wastes compute when it cannot identify which agent's intermediate message advanced progre

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation

Model ReleasesDGX agent

arXiv:2607.09142v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in online medical consultation, yet existing benchmarks remain poorly aligned with real clini

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

SafetyDGX agent

arXiv:2607.13591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly rely on external memory systems to accumulate experience across tasks. Yet nearly all existing approach

Mind the Gap: Action Rebinding Attacks against Android GUI Agents

SafetyDGX agent

arXiv:2601.12349v3 Announce Type: replace-cross Abstract: Large multimodal model powered GUI agents are emerging as high-privilege operators on mobile platforms, entrusted to perceive screen content a

Multi-Agent Collaborative Reasoning with Tool-Augmented Evidence for Urban Region Profiling

AgentsDGX agent

arXiv:2607.13558v1 Announce Type: new Abstract: Urban region profiling constitutes a core problem in urban computing, supporting applications such as population estimation, economic assessment, and en

Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study

ApplicationsDGX agent

arXiv:2607.14041v1 Announce Type: cross Abstract: Historical Manchu OCR must accommodate various visually distinct writing styles, including regular script, running script, and the semi-cursive chance

Multimodal Assessment of Pancreatic Cancer Resectability Using Deep Learning

Local AiDGX agent

arXiv:2607.13826v1 Announce Type: cross Abstract: Accurate determination of pancreatic ductal adenocarcinoma (PDAC) resectability relies on evaluating how the tumor interacts with major peripancreatic

Music-to-Dance Generation via Atomic Movements

SafetyDGX agent

arXiv:2607.13978v1 Announce Type: cross Abstract: Music-driven dance generation aims to produce human motion that is both rhythmically synchronized and semantically consistent with music. While recent

MxGPS: Multiplex Graph Transformers for a Power Grid Foundation Model

Model ReleasesDGX agent

arXiv:2607.13763v1 Announce Type: cross Abstract: Single-task fine-tuning of graph neural networks (GNNs) for power grid problems exhibits a systematic failure mode: models that achieve the lowest in-

Networked Intelligence: Active Shared Context Graphs for Human-AI Team Science

AgentsDGX agent

arXiv:2607.13220v1 Announce Type: new Abstract: Most AI-for-science systems focus on scaling a single reasoning process through better models, larger context windows, long-horizon agentic execution, o

NodeImport: Imbalanced Node Classification with Node Importance Assessment

ApplicationsDGX agent

arXiv:2607.13837v1 Announce Type: cross Abstract: In real-world applications, node classification on graphs often faces the challenge of class imbalance, where majority classes dominate training, resu

← Previous
1…6364656667…354
Next →