AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation

DGX agent

arXiv:2607.09142v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in online medical consultation, yet existing benchmarks remain poorly aligned with real clini

model-releasesarxiv-cs-ai
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

DGX agent

arXiv:2607.13591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly rely on external memory systems to accumulate experience across tasks. Yet nearly all existing approach

safetyarxiv-cs-ai
16 Jul 2026
Safety

Mind the Gap: Action Rebinding Attacks against Android GUI Agents

DGX agent

arXiv:2601.12349v3 Announce Type: replace-cross Abstract: Large multimodal model powered GUI agents are emerging as high-privilege operators on mobile platforms, entrusted to perceive screen content a

safetyarxiv-cs-ai
16 Jul 2026
Agents

Multi-Agent Collaborative Reasoning with Tool-Augmented Evidence for Urban Region Profiling

DGX agent

arXiv:2607.13558v1 Announce Type: new Abstract: Urban region profiling constitutes a core problem in urban computing, supporting applications such as population estimation, economic assessment, and en

agentsarxiv-cs-ai
16 Jul 2026
Applications

Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study

DGX agent

arXiv:2607.14041v1 Announce Type: cross Abstract: Historical Manchu OCR must accommodate various visually distinct writing styles, including regular script, running script, and the semi-cursive chance

applicationsarxiv-cs-ai
16 Jul 2026
Local Ai

Multimodal Assessment of Pancreatic Cancer Resectability Using Deep Learning

DGX agent

arXiv:2607.13826v1 Announce Type: cross Abstract: Accurate determination of pancreatic ductal adenocarcinoma (PDAC) resectability relies on evaluating how the tumor interacts with major peripancreatic

local-aiarxiv-cs-ai
16 Jul 2026
Safety

Music-to-Dance Generation via Atomic Movements

DGX agent

arXiv:2607.13978v1 Announce Type: cross Abstract: Music-driven dance generation aims to produce human motion that is both rhythmically synchronized and semantically consistent with music. While recent

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

MxGPS: Multiplex Graph Transformers for a Power Grid Foundation Model

DGX agent

arXiv:2607.13763v1 Announce Type: cross Abstract: Single-task fine-tuning of graph neural networks (GNNs) for power grid problems exhibits a systematic failure mode: models that achieve the lowest in-

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Networked Intelligence: Active Shared Context Graphs for Human-AI Team Science

DGX agent

arXiv:2607.13220v1 Announce Type: new Abstract: Most AI-for-science systems focus on scaling a single reasoning process through better models, larger context windows, long-horizon agentic execution, o

agentsarxiv-cs-ai
16 Jul 2026
Applications

NodeImport: Imbalanced Node Classification with Node Importance Assessment

DGX agent

arXiv:2607.13837v1 Announce Type: cross Abstract: In real-world applications, node classification on graphs often faces the challenge of class imbalance, where majority classes dominate training, resu

applicationsarxiv-cs-ai
16 Jul 2026
Model Releases

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs

DGX agent

arXiv:2601.02023v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly utilize massive context windows as working memory for autonomous tasks, their reliability fluctua

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache

DGX agent

arXiv:2505.18231v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference is typically memory-intensive, especially when processing large batch sizes and long sequences, due to th

safetyarxiv-cs-ai
16 Jul 2026
Safety

Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows

DGX agent

arXiv:2607.13078v1 Announce Type: cross Abstract: LLMs are now proposed for fraud detection, scam investigation, content moderation, and other trust-and-safety workflows. Much of the public literature

safetyarxiv-cs-ai
16 Jul 2026
Agents

Oracle Agent Memory as an Enterprise Memory Substrate for Long-Horizon AI Agents

DGX agent

arXiv:2607.13157v1 Announce Type: new Abstract: Agent memory is a systems problem for long-horizon agents. Practical deployments require retention of task state across extended conversations, recovery

agentsarxiv-cs-ai
16 Jul 2026
Research

OriginBlame: Record- and Token-Level Data Provenance for AI Training Datasets

DGX agent

arXiv:2607.13037v1 Announce Type: new Abstract: When a data contributor requests removal, model trainers face a practical gap: unlearning algorithms require a forget set, yet no tool can locate which

researcharxiv-cs-ai
16 Jul 2026
Model Releases

OvisOCR2 Technical Report

DGX agent

arXiv:2607.13639v1 Announce Type: cross Abstract: We introduce OvisOCR2, a 0.8B document parsing model. OvisOCR2 is designed as an end-to-end parser: given a document page image, it generates a Markdo

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Partially Correlated Verifier Cascades in LLM Harnesses: Concave Log-Odds, Polynomial Reliability, and Blind-Spot Ceilings

DGX agent

arXiv:2607.13918v1 Announce Type: cross Abstract: Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if k verifier calls all accept it. Un

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

PC-Diffuser: Path-Consistent Capsule CBF Safety Filtering for Diffusion-Based Trajectory Planner

DGX agent

arXiv:2603.10330v2 Announce Type: replace-cross Abstract: Autonomous driving in complex traffic requires planners that generalize beyond hand-crafted rules, motivating data-driven approaches that lear

safetyarxiv-cs-ai
16 Jul 2026
Research

PersGuard: Preventing Malicious Personalization in Text-to-Image Diffusion Models via Model Backdoors

DGX agent

arXiv:2502.16167v2 Announce Type: replace-cross Abstract: Diffusion models (DMs) have advanced text-to-image (T2I) synthesis, yet their personalization capabilities raise serious privacy and copyright

researcharxiv-cs-ai
16 Jul 2026
Model Releases

Policy of Thoughts: Scaling Test-Time Training for LLM Reasoning via Online Policy Evolution

DGX agent

arXiv:2601.20379v2 Announce Type: replace Abstract: Large language models (LLMs) struggle with complex, long-horizon reasoning due to instability caused by their frozen policy assumption. Current test

model-releasesarxiv-cs-ai
16 Jul 2026
Research

Post-Disaster Affected Area Segmentation with a Vision Transformer (ViT)-based EVAP Model using Sentinel-2 and Formosat-5 Imagery

DGX agent

arXiv:2507.16849v3 Announce Type: replace-cross Abstract: We propose a vision transformer (ViT)-based deep learning framework to refine disaster-affected area segmentation from remote sensing imagery,

researcharxiv-cs-ai
16 Jul 2026
Safety

Price of Fairness in Bandits: A Tight Minimax Characterization

DGX agent

arXiv:2607.13402v1 Announce Type: cross Abstract: In bandit problems, standard regret-minimizing algorithms treat exploration as an amortized cost, which can expose early participants to unfair ex-ant

safetyarxiv-cs-ai
16 Jul 2026
Safety

Privacy Preserving Recommender Systems Balancing Personalization with Privacy

DGX agent

arXiv:2607.13328v1 Announce Type: cross Abstract: Personalized recommendation systems are central to modern e-commerce and retail platforms, but they typically rely on centralized storage of detailed

safetyarxiv-cs-ai
16 Jul 2026
Research

Probabilistic Extension of Neuro-Symbolic AGI Robots based on Belnap's Typed Intensional FOL

DGX agent

arXiv:2607.13073v1 Announce Type: new Abstract: Neuro-symbolic AI based on IFOL_B is a way to combine neural learning and symbolic reasoning to overcome limitations of purely neural systems (like lack

researcharxiv-cs-ai
16 Jul 2026
Safety

Protective Capacity Hallucination: When Large Language Models Claim Nonexistent Capabilities

DGX agent

arXiv:2607.13596v1 Announce Type: cross Abstract: When cast as the protector of a vulnerable user yet given no explicit capability boundary, a large language model (LLM) may respond not by acknowledgi

safetyarxiv-cs-ai
16 Jul 2026
Safety

RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset

DGX agent

arXiv:2603.11811v2 Announce Type: replace-cross Abstract: The acquisition of large-scale physical interaction data, a critical prerequisite for modern robot learning, is severely bottlenecked by the p

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

DGX agent

arXiv:2607.13189v1 Announce Type: cross Abstract: We present RAGthoven, our system for SemEval-2026 Task 1 (MWAHAHA), Subtask A (multilingual constrained humor generation in English, Spanish, and Chin

model-releasesarxiv-cs-ai
16 Jul 2026
Research

Reassessing Muon for Matrix Factorization

DGX agent

arXiv:2607.13246v1 Announce Type: cross Abstract: Muon has recently emerged as a strong optimizer for large-scale deep learning, where it reshapes gradient updates through approximate orthogonalizatio

researcharxiv-cs-ai
16 Jul 2026
Model Releases

Representation-Based Exploration for Language Models: From Test-Time to Post-Training

DGX agent

arXiv:2510.11686v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to expand the capabilities of language models, but it is unclear if current RL techniques promote the dis

model-releasesarxiv-cs-ai
16 Jul 2026
Research

Rethinking Multimodal Fusion for Time Series: Text Modalities Need Constrained Fusion

DGX agent

arXiv:2603.22372v3 Announce Type: replace-cross Abstract: Recent advances in multimodal learning have motivated the integration of auxiliary modalities such as text or vision into time series (TS) for

researcharxiv-cs-ai
16 Jul 2026
Agents

Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavioral Objective Violation

DGX agent

arXiv:2607.14006v1 Announce Type: cross Abstract: Penetration testing traditionally evaluates whether adversaries can exploit weaknesses in software, infrastructure, configurations, or operational con

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Safeguard-Conditioned Uplift: Measuring Utility-Risk Frontiers for Dual-Use Biology Assistants

DGX agent

arXiv:2607.13039v1 Announce Type: cross Abstract: Safety evaluations for dual-use biology assistants often measure base-model capability, refusal behavior, or jailbreak success. These metrics miss a d

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing

DGX agent

arXiv:2607.13594v1 Announce Type: new Abstract: LLM agents act on real-world environments through tool calls, and a single misjudged action can cause irreversible harm. The standard safeguard is a gua

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

ScanFocus: A Coarse-to-Fine Framework for Spatio-Temporal Video Grounding

DGX agent

arXiv:2607.13421v1 Announce Type: cross Abstract: Spatio-Temporal Video Grounding (STVG) aims to retrieve the visual trajectory of a specific object from a video stream as described by a natural langu

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Self-Improvements in Modern Agentic Systems: A Survey

DGX agent

arXiv:2607.13104v1 Announce Type: new Abstract: Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is controllable evolution, or adaptation, fro

agentsarxiv-cs-ai
16 Jul 2026
Agents

Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Framework

DGX agent

arXiv:2607.13091v1 Announce Type: cross Abstract: LLM-based coding agents repeat the same classes of mistakes across sessions because they lack a mechanism to retain corrections from human review feed

agentsarxiv-cs-ai
16 Jul 2026
Local Ai

SemaDiff: Identifying Semantic-Changing Commits with Generated Code and Tests

DGX agent

arXiv:2607.13111v1 Announce Type: cross Abstract: Distinguishing semantic-preserving commits from changing ones remains an open challenge in software repository mining. While existing approaches detec

local-aiarxiv-cs-ai
16 Jul 2026
Applications

Semantic Anchoring for Robotic Action Representations

DGX agent

arXiv:2607.13597v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models inherit rich semantic representations from pretrained Vision-Language Models, yet fine-tuning on limited robot dem

applicationsarxiv-cs-ai
16 Jul 2026
Model Releases

Set-shifting Behavioral Test for Harnessed Agents

DGX agent

arXiv:2607.13396v1 Announce Type: new Abstract: What happens to an LLM agent's tool choice when the reliable tool silently changes within an ongoing session? We borrow set-shifting from cognitive psyc

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation

DGX agent

arXiv:2607.13124v1 Announce Type: cross Abstract: Structured pruning is a hardware-friendly way to compress LLMs, but it is mostly validated on multiple-choice recognition tasks, while the same compre

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification

DGX agent

arXiv:2607.13081v1 Announce Type: cross Abstract: We present nsfaguard, a guardrail framework for securing agentic AI systems against operational threats, such as prompt injection, sensitive informati

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Social Simulations: from Agent-Based Modeling to Digital Twins

DGX agent

arXiv:2607.13693v1 Announce Type: cross Abstract: This book chapter covers the evolution of social simulation from classical agent-based models, in which agents interact according to explicitly define

agentsarxiv-cs-ai
16 Jul 2026
Research

Spectral-Informed Neural Networks Outperform Spectral Methods in High-dimensional PDEs

DGX agent

arXiv:2607.13566v1 Announce Type: cross Abstract: For low-dimensional problems (dleq3), spectral methods can achieve exceptionally high accuracy. For middle-dimensional problems (4 leq d lesssim 10),

researcharxiv-cs-ai
16 Jul 2026
Model Releases

SPINE: Bridging the Cyber-Physical Gap with Agentic AI

DGX agent

arXiv:2607.13049v1 Announce Type: new Abstract: Foundation models have given robots a sophisticated brain for complex decision-making, yet deploying that intelligence into a physical platform still de

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy

DGX agent

arXiv:2607.13175v1 Announce Type: cross Abstract: Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a criterion that often fails to detect rare but catastrop

safetyarxiv-cs-ai
16 Jul 2026
Applications

STKAN: Kolmogorov-Arnold Networks for Spatio-Temporal Forecasting

DGX agent

arXiv:2607.13108v1 Announce Type: cross Abstract: Real-world traffic data exhibit heterogeneous spatial correlations and nonlinear temporal dynamics, posing substantial challenges for accurate spatio-

applicationsarxiv-cs-ai
16 Jul 2026
Model Releases

STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle

DGX agent

arXiv:2607.13618v1 Announce Type: new Abstract: LLM agents are increasingly evaluated on multi-week decision tasks in which the state that drives cost is never directly observed. On such tasks the fin

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Symbiosis-Inspired Knowledge Distillation for Incremental Object Detection

DGX agent

arXiv:2607.13452v1 Announce Type: cross Abstract: Incremental object detection (IOD) aims to extend detectors to new categories while retaining previously acquired knowledge. Existing methods often ad

model-releasesarxiv-cs-ai
16 Jul 2026
← Previous
1…8586878889…448
Next →