AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Population-Based Multi-Objective Training of Discriminators for Semi-Supervised GANs

DGX agent

arXiv:2607.01907v1 Announce Type: cross Abstract: Semi-supervised generative adversarial networks (SSL-GANs) can exploit large unlabeled datasets while retaining a classifier in the discriminator, but

researcharxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Power Systems Agent Benchmark: Executable Evaluation of AI Agents in Electric Power Engineering

DGX agent

arXiv:2606.20950v2 Announce Type: replace Abstract: Executable evaluation -- checking the consequences of an agent's actions with a program rather than grading its prose -- has become a prominent way

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

PPTArena: A Benchmark for PowerPoint Editing

DGX agent

arXiv:2512.03042v3 Announce Type: replace-cross Abstract: We introduce PPTArena, a benchmark for PowerPoint editing that evaluates how agents modify real slides from natural-language instructions. Unl

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Pre-Flight: A Benchmark for Evaluating Large Language Models on Aviation Operational Knowledge

DGX agent

arXiv:2607.01829v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed for aviation business operations, from documentation and training generation to customer facing a

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

DGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

safetyarxiv-cs-ai
3 Jul 2026
Research

Predicting Early Stages Of Alzheimer's Disease And Identifying Key Biomarkers Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies

DGX agent

arXiv:2607.02142v1 Announce Type: cross Abstract: Alzheimers disease (AD) is a brain disorder that develops slowly and mainly affects memory, thinking, language, and daily activities. It is one of the

researcharxiv-cs-ai
3 Jul 2026
Model Releases

PreScience: A Dataset and Benchmark for Scientific Forecasting

DGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

model-releasesarxiv-cs-ai
3 Jul 2026
Local Ai

ProCal: Inference-Time Proposal Calibration for Open-Vocabulary Object Detection

DGX agent

arXiv:2607.01759v1 Announce Type: cross Abstract: Open-vocabulary object detection aims to localize and classify objects beyond the fixed set of categories seen dur ing training. Recent open-vocabular

local-aiarxiv-cs-ai
3 Jul 2026
Local Ai

Procedural Memory Distillation: Online Reflection for Self-Improving Language Models

DGX agent

arXiv:2607.01480v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), along with recent selfdistillation variants such as SDPO, evaluates each rollout against a verifi

local-aiarxiv-cs-ai
3 Jul 2026
Applications

Profit-Based Counterfactual Explanations for Product Improvement: A Case Study of Manga Sales in Japan

DGX agent

arXiv:2607.01610v1 Announce Type: new Abstract: Counterfactual explanation (CE) is widely used to enhance the interpretability of machine learning models and support data-driven decision-making based

applicationsarxiv-cs-ai
3 Jul 2026
Model Releases

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

DGX agent

arXiv:2607.02512v1 Announce Type: cross Abstract: Many everyday programming tasks resist clean rule-based implementation, such as alerting on important log lines, repairing malformed JSON, or ranking

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Prompt Coverage Adequacy

DGX agent

arXiv:2607.02057v1 Announce Type: cross Abstract: In recent years, it has become increasingly evident that large language models (LLMs) and autonomous agents raise the level of abstraction in software

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Prompt Framing Distorts Count-Based Evaluation of LLM Error Detection: Evidence from Numeric Anchoring

DGX agent

arXiv:2607.01240v1 Announce Type: cross Abstract: Count-based F1 is widely used as a proxy for LLM error-detection quality, but this paper shows that it can rise dramatically without a corresponding i

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Psychological Imagination Networks Show Cross-Population Centrality and Clustering Alignment in Humans That Large Language Models Fail to Replicate

DGX agent

arXiv:2510.04391v5 Announce Type: replace Abstract: Mental imagery vividness is a stable individual trait, yet whether imagined scenarios share relational structure across human and synthetic large la

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness

DGX agent

arXiv:2510.04484v2 Announce Type: replace-cross Abstract: The ability to control LLMs' emulated emotional states and personality traits is an essential step in enabling rich, human-centered interactio

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

DGX agent

arXiv:2607.02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference

safetyarxiv-cs-ai
3 Jul 2026
Model Releases

QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition

DGX agent

arXiv:2607.02426v1 Announce Type: cross Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

RadiomicNet: A Hybrid Radiomics-Guided Lightweight Architecture for Interpretable Medical Image Segmentation

DGX agent

arXiv:2607.02185v1 Announce Type: cross Abstract: Deep learning has achieved remarkable performance in medical image segmentation, yet it suffers from critical limitations: mathematical intractability

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Rank-Then-Act: Reward-Free Control from Frame-Order Progress

DGX agent

arXiv:2607.01897v1 Announce Type: cross Abstract: We introduce Rank-Then-Act (RTA), a framework for learning control policies from expert video demonstrations without environment rewards. RTA trains a

safetyarxiv-cs-ai
3 Jul 2026
Local Ai

Reasoning effort, not tool access, buys first-try reliability in agentic code generation: an observational study

DGX agent

arXiv:2607.02436v1 Announce Type: cross Abstract: Agentic coding assistants are increasingly given extra capabilities, such as browser based testing tools and design oriented system prompts, on the as

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Reasoning LLM Improves Speaker Recognition in Long-form TV Dramas

DGX agent

arXiv:2607.02504v1 Announce Type: cross Abstract: Long-form TV dramas present a formidable challenge for comprehensive video understanding, where deciphering complex storyline often relies on extbf{sp

model-releasesarxiv-cs-ai
3 Jul 2026
Research

ReContext: Recursive Evidence Replay as LLM Harness for Long-Context Reasoning

DGX agent

arXiv:2607.02509v1 Announce Type: new Abstract: Understanding and reasoning over long contexts has become a key requirement for deploying large language models (LLMs) in realistic applications. Althou

researcharxiv-cs-ai
3 Jul 2026
Safety

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

DGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

safetyarxiv-cs-ai
3 Jul 2026
Applications

Reformalization of the Jordan Curve Theorem

DGX agent

arXiv:2607.01734v1 Announce Type: new Abstract: We present a case study in reformalization, a variant of autoformalization in which the input proof is not natural language but a formal development in

applicationsarxiv-cs-ai
3 Jul 2026
Local Ai

Repair the Amplifier, Not the Symptom: Stable World-Model Correction for Agent Rollouts

DGX agent

arXiv:2607.01767v1 Announce Type: new Abstract: As agent planning moves from short tool chains toward persistent workflows with thousands or tens of thousands of steps, failures will occur inside larg

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration

DGX agent

arXiv:2603.06001v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasing

model-releasesarxiv-cs-ai
3 Jul 2026
Research

Rethinking Complexity Metrics for LLM-Integrated Applications: Beyond Source Code

DGX agent

arXiv:2607.01903v1 Announce Type: new Abstract: LLM-integrated applications blend natural language prompts with program code, and much of their runtime behavior originates in the prompt layer rather t

researcharxiv-cs-ai
3 Jul 2026
Local Ai

Rethinking Generic Object Tracking Toward Human-Level Perceptual Intelligence

DGX agent

arXiv:2607.01395v1 Announce Type: cross Abstract: At the heart of human visual perception lies the ability to maintain a continuous and coherent understanding of the external world. By integrating obs

local-aiarxiv-cs-ai
3 Jul 2026
Research

Revisiting Chain-of-Thought Reasoning under Limited Supervision: Semi-supervised Chain-of-Thought Learning

DGX agent

arXiv:2607.01511v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has emerged as an effective approach for activating latent reasoning capabilities in large language models. However, mo

researcharxiv-cs-ai
3 Jul 2026
Safety

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

DGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

safetyarxiv-cs-ai
3 Jul 2026
Research

Robust and Explainable 3D Mode Shape Recognition Using Region-Aware Graph Neural Networks

DGX agent

arXiv:2607.01522v1 Announce Type: cross Abstract: Mode shape recognition is a fundamental task in automotive NVH development, yet it remains dependent on manual visual inspection by experienced engine

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Robust for the Wrong Reasons: The Representational Geometry of LLM Robustness to Science Skepticism

DGX agent

arXiv:2607.01951v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly consulted on contested scientific questions, raising the concern that they will sycophantically retreat

model-releasesarxiv-cs-ai
3 Jul 2026
Research

SA-HGNN: Sample-Adaptive Hyperbolic Graph Neural Network for EEG-Based Depression Recognition

DGX agent

arXiv:2607.02063v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have been widely used to capture spatial functional connectivity patterns to improve electroencephalography (EEG)-based d

researcharxiv-cs-ai
3 Jul 2026
Model Releases

SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models

DGX agent

arXiv:2607.01876v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in multimodal understanding, yet their enormous parameter scale and cross-modal

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

SABER: A Semantic-Aligned Brain Network Analysis Framework via Multi-scale Hypergraphs

DGX agent

arXiv:2607.01901v1 Announce Type: cross Abstract: Effective brain disease diagnosis requires the synergy of brain connectivity patterns and high-level semantic knowledge. Existing methods, however, la

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

DGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safeguarding LLM Agents from Misalignment through Provenance Analysis

DGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

safetyarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Targeted Embedding Exploit via Refinement

DGX agent

arXiv:2607.01859v1 Announce Type: new Abstract: Safety training for large language models (LLMs) is conducted predominantly in English, leaving uncertain how well safety mechanisms generalize to low-r

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

DGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

model-releasesarxiv-cs-ai
3 Jul 2026
Tutorials

Scaling Laws for Grid-Based Approximate Nearest Neighbor Search in High Dimensions

DGX agent

arXiv:2607.01283v1 Announce Type: cross Abstract: Grid-based approaches to approximate nearest neighbor (ANN) search have been absent from modern scaling analyses. We present a systematic characteriza

tutorialsarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling Trends for Lie Detector Oversight in Preference Learning

DGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling

DGX agent

arXiv:2607.01612v1 Announce Type: new Abstract: Training large language models (LLMs) with reinforcement learning (RL) has significantly advanced their performance on reasoning and question-answering

model-releasesarxiv-cs-ai
3 Jul 2026
Research

Scene-Conditioned PINN-GNN for Multipath RF Maps: Cross-Scene Generation and In-Scene Completion

DGX agent

arXiv:2607.01777v1 Announce Type: cross Abstract: Radio frequency (RF) maps provide a compact representation of multipath propagation characteristics and are fundamental to channel modeling, coverage

researcharxiv-cs-ai
3 Jul 2026
Tutorials

SelectTSL: Prompt-Guided Selective Target Sound Localization in Complex Scenarios

DGX agent

arXiv:2607.02343v1 Announce Type: cross Abstract: Humans can selectively attend to a target sound and estimate its direction in complex scenarios, whereas such selective localization remains challengi

tutorialsarxiv-cs-ai
3 Jul 2026
Model Releases

Self-Gating Attention for Efficient Time Series Forecasting

DGX agent

arXiv:2607.02344v1 Announce Type: cross Abstract: Transformer architectures have shown strong potential in time series forecasting, where multi-head self-attention is widely used to capture temporal d

model-releasesarxiv-cs-ai
3 Jul 2026
Research

SemHash-LLM: A Multi-Granularity Semantic Hashing Framework for Document Deduplication

DGX agent

arXiv:2607.01601v1 Announce Type: new Abstract: Large scale document deduplication must preserve semantic equivalence while remaining efficient over massive corpora. We present SemHash LLM, a multi gr

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

DGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

model-releasesarxiv-cs-ai
3 Jul 2026
Local Ai

SEPS: Semantic-enhanced Patch Slimming Framework for fine-grained cross-modal alignment

DGX agent

arXiv:2511.01390v2 Announce Type: replace-cross Abstract: Fine-grained cross-modal alignment aims to establish precise local correspondences between vision and language, forming a cornerstone for visu

local-aiarxiv-cs-ai
3 Jul 2026
← Previous
1…121122123124125…448
Next →