AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

On the Separability of Information in Diffusion Models

DGX agent

arXiv:2509.23937v5 Announce Type: replace-cross Abstract: Diffusion models transform noise into data by injecting information that was captured in their neural network during the training phase. In th

researcharxiv-cs-ai
23 Jul 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On the Systematic Challenges of Culturally Loaded Machine Translation: Dream of the Red Chamber as the Cultural Lens

DGX agent

arXiv:2607.20241v1 Announce Type: cross Abstract: Culturally loaded translation poses unique challenges for machine translation (MT), as meanings are deeply embedded in socio-cultural contexts beyond

researcharxiv-cs-ai
23 Jul 2026
Safety

OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks

DGX agent

arXiv:2607.19351v1 Announce Type: new Abstract: LLM-based multi-agent systems (LLM-MAS) are increasingly deployed in safety-critical applications, where adversaries inject malicious instructions throu

safetyarxiv-cs-ai
23 Jul 2026
Safety

OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization

DGX agent

arXiv:2607.19806v1 Announce Type: cross Abstract: Activation steering provides a lightweight mechanism for controlling large language models at inference time, but steering vectors can have unintended

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Opto-ViT-v2: Noise-Resilient On-Chip Fine-Tuning for Photonic Near-Sensor Vision Transformer Accelerators

DGX agent

arXiv:2607.19421v1 Announce Type: cross Abstract: Silicon-photonic (SiPh) accelerators have emerged as a promising platform for Vision Transformer (ViT) inference by performing matrix multiplications

model-releasesarxiv-cs-ai
23 Jul 2026
Research

OSVE: One Step Video Editing with One Step Diffusion Models

DGX agent

arXiv:2607.19895v1 Announce Type: cross Abstract: Text-guided video editing with diffusion models is impractically slow, hindered by costly multi-step sampling and inversion. We present OSVE, the firs

researcharxiv-cs-ai
23 Jul 2026
Applications

Overview of FinMMEval 2026 Task 1: Multilingual Financial Multiple-Choice Question Answering

DGX agent

arXiv:2607.19856v1 Announce Type: cross Abstract: FinMMEval 2026 Task 1 evaluates multilingual financial multiple-choice question answering in English, Chinese, Arabic, and Hindi. The task tests wheth

applicationsarxiv-cs-ai
23 Jul 2026
Research

Overview of FinMMEval 2026 Task 2: Multilingual Financial Short-Answer Question Answering

DGX agent

arXiv:2607.19867v1 Announce Type: cross Abstract: FinMMEval 2026 Task 2 evaluates short-answer financial question answering over multilingual evidence. Each final-test item pairs an English question w

researcharxiv-cs-ai
23 Jul 2026
Model Releases

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization

DGX agent

arXiv:2607.19653v1 Announce Type: cross Abstract: Large language model (LLM) agents now perform well on correctness-oriented repository-level tasks, including SWE-Bench issue resolution and feature im

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Persian Pixel: A large-scale synthetic OCR dataset for Persian language

DGX agent

arXiv:2607.20385v1 Announce Type: cross Abstract: Optical Character Recognition (OCR) for Persian remains substantially less mature than for Latin-script languages despite Persian being spoken by more

researcharxiv-cs-ai
23 Jul 2026
Agents

Personalized Recommendation Tool Learning via Autonomous Language Agents

DGX agent

arXiv:2607.19739v1 Announce Type: cross Abstract: Although large language models (LLMs) have recently gained traction in recommender systems due to their strong reasoning capabilities and extensive wo

agentsarxiv-cs-ai
23 Jul 2026
Safety

PGTT: Phase-Guided Terrain Traversal for Perceptive Legged Locomotion

DGX agent

arXiv:2510.18348v2 Announce Type: replace-cross Abstract: State-of-the-art perceptive Reinforcement Learning controllers for legged robots typically either (i) impose oscillator-or IK-based gait prior

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

PhenSPINE: A Standardized Benchmark for Spine Pathology Diagnosis

DGX agent

arXiv:2607.19696v1 Announce Type: cross Abstract: The accurate diagnosis of spinal pathologies depends heavily on radiological interpretation, yet automated systems are hindered by the lack of diverse

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Physics-Aware Complex-Valued State Space Model with Scattering-Prior Feature Modulation for PolSAR Image Classification

DGX agent

arXiv:2607.19787v1 Announce Type: cross Abstract: Polarimetric synthetic aperture radar (PolSAR) image classification is a representative task for physics-aware GeoAI, where land-cover semantics are c

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Post-Training in Time Series Foundation Models: A Unifying Framework

DGX agent

arXiv:2607.20002v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have emerged as general-purpose models for time series analysis, but pretraining alone is often insufficient for

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity

DGX agent

arXiv:2607.20268v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at many tasks, they frequently struggle with complex reasoning that requires long-horizon planning and iterativ

agentsarxiv-cs-ai
23 Jul 2026
Local Ai

Pre-Deployment Complexity Estimation for Federated Perception Systems

DGX agent

arXiv:2603.28282v2 Announce Type: replace-cross Abstract: Edge AI systems increasingly rely on federated learning to train perception models in distributed, privacy-preserving, and resource-constraine

local-aiarxiv-cs-ai
23 Jul 2026
Safety

Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models

DGX agent

arXiv:2607.19453v1 Announce Type: cross Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance S

safetyarxiv-cs-ai
23 Jul 2026
Research

PRIME-SVR: Physics-infoRmed Implicit Multi-Echo Slice-to-Volume Reconstruction for Fetal T2 mapping

DGX agent

arXiv:2607.20136v1 Announce Type: cross Abstract: Slice-to-volume reconstruction (SVR) is the standard method for obtaining high-resolution (HR) 3D fetal brain volumes from motion-corrupted 2D MRI sli

researcharxiv-cs-ai
23 Jul 2026
Research

PRISM-DR: Per-lesion Retinal Inference with Specialist Models for Diabetic Retinopathy

DGX agent

arXiv:2607.19864v1 Announce Type: cross Abstract: Diabetic retinopathy is a leading cause of preventable blindness; its early lesions are small, low contrast, and easily missed in manual screening. Mo

researcharxiv-cs-ai
23 Jul 2026
Agents

PRO-LONG: Programmatic Memory Enables Long-Horizon Reasoning

DGX agent

arXiv:2607.20064v1 Announce Type: new Abstract: Long-horizon tasks require sustained perception, reasoning, and exploration, and are a persistent challenge for large language model (LLM) agents. This

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

Prober.ai: Gated Inquiry-Based Feedback via LLM-Constrained Personas for Argumentative Writing Development

DGX agent

arXiv:2605.05598v2 Announce Type: replace Abstract: The proliferation of large language models (LLMs) in educational settings has paradoxically undermined the cognitive processes they purport to suppo

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Profile-Graph Memory for LLM Agents: Implicit Cross-Entity Traversal through Narrative Profiles

DGX agent

arXiv:2607.19359v1 Announce Type: new Abstract: Long-term memory is essential for LLM agents that interact across sessions, yet current memory benchmarks primarily evaluate single-hop recall, leaving

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Prompt Programming for Cultural Bias and Alignment of Large Language Models

DGX agent

arXiv:2603.16827v2 Announce Type: replace Abstract: Culture shapes reasoning, values, prioritization, and strategic decision-making, yet large language models (LLMs) often exhibit cultural biases that

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering

DGX agent

arXiv:2607.20253v1 Announce Type: cross Abstract: In this report, we present a unified song generation framework capable of producing high-quality full-length music from lyrics, text descriptions, and

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Rater State Bias in RLHF Preference Data: An Audit Framework

DGX agent

arXiv:2607.16195v2 Announce Type: replace Abstract: We identify a structured confound in Reinforcement Learning from Human Feedback (RLHF). Pairwise preference labels are intended to reflect the compa

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model

DGX agent

arXiv:2607.20058v1 Announce Type: new Abstract: Large language models can answer scientific questions, yet a correct output does not reveal whether the model represents or uses the governing physics.

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Recovering Clinical Utility Under Differential Privacy: Empirical Validation of Adaptive Federated Aggregation on Heterogeneous Cardiovascular Datasets

DGX agent

arXiv:2607.19403v1 Announce Type: cross Abstract: Validating federated learning frameworks on real clinical data is an essential step between proof-of-concept demonstrations in controlled synthetic en

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Reference-Free Evaluation of Reasoning in Open-Ended Question Answering

DGX agent

arXiv:2607.19678v1 Announce Type: cross Abstract: AI-generated answers in high-stakes domains are often fluent but difficult to verify, especially when they contain multi-step reasoning rather than a

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning

DGX agent

arXiv:2607.19450v1 Announce Type: cross Abstract: Large-scale online reinforcement learning (RL) is the predominant means of eliciting advanced abilities including long-term reasoning and agentic tool

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results

DGX agent

arXiv:2607.20090v1 Announce Type: cross Abstract: Retrieval-augmented large language models frequently face contexts that interleave useful evidence with misleading statements or instruction-like cont

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Rethinking Uncertainty Evaluation in Large Language Models

DGX agent

arXiv:2607.19367v1 Announce Type: new Abstract: Calibration is the primary criterion for evaluating LLM confidence, but it is insufficient: it admits trivially incoherent estimators, depends on the ev

researcharxiv-cs-ai
23 Jul 2026
Safety

Rewarding Better Thinking for LLM Preference Alignment

DGX agent

arXiv:2607.19824v1 Announce Type: new Abstract: LLM preference alignment aims to optimize models toward human preferences across diverse user instructions. Reinforcement learning has become a major po

safetyarxiv-cs-ai
23 Jul 2026
Research

RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling

DGX agent

arXiv:2607.19776v1 Announce Type: cross Abstract: Existing symbolic music generation models typically use bars as the basic structural unit. However, human perception of musical phrases often does not

researcharxiv-cs-ai
23 Jul 2026
Model Releases

Safe Remediation as Risk-Constrained Intervention Decision in Microservice Systems

DGX agent

arXiv:2607.20005v1 Announce Type: new Abstract: In modern IT operations (IT-Ops), the cost of an incorrect repair often exceeds the cost of no action at all. Yet existing automated remediation systems

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses

DGX agent

arXiv:2607.19387v1 Announce Type: cross Abstract: Surrogate modeling for high-dimensional nonlinear dynamical systems that exhibit chaos requires mechanisms that preserve not only pointwise accuracy b

safetyarxiv-cs-ai
23 Jul 2026
Hardware

Scaling Time Series Classification via XAI-Driven Data Reduction

DGX agent

arXiv:2607.15774v2 Announce Type: replace-cross Abstract: Explainable AI (XAI) for time series has seen significant algorithmic growth, but its utility in providing measurable performance gains for do

hardwarearxiv-cs-ai
23 Jul 2026
Applications

Schrodinger Bridge Mamba for One-Step Speech Enhancement

DGX agent

arXiv:2510.16834v3 Announce Type: replace-cross Abstract: We present Schrodinger Bridge Mamba (SBM), a novel model for efficient speech enhancement by integrating the Schrodinger Bridge (SB) training

applicationsarxiv-cs-ai
23 Jul 2026
Research

SciTrek: Evaluating and Improving Long-Context Numerical Reasoning over Scientific Articles

DGX agent

arXiv:2509.21028v4 Announce Type: replace Abstract: We introduce SciTrek, a synthetic question-answering dataset for assessing and improving long-context numerical reasoning in large language models (

researcharxiv-cs-ai
23 Jul 2026
Tutorials

SCPP: A Unified Python Library for Soft Clustering

DGX agent

arXiv:2607.19620v1 Announce Type: cross Abstract: In this paper, we present SCPP (Soft Clustering Python Package), an open-source Python framework for soft clustering. SCPP establishes a canonical, sc

tutorialsarxiv-cs-ai
23 Jul 2026
Local Ai

Self-supervision drives representational convergence in medical foundation models more than clinical supervision

DGX agent

arXiv:2607.20274v1 Announce Type: cross Abstract: Medical image encoders from different groups are increasingly treated as interchangeable, on the assumption that scale and clinical supervision concen

local-aiarxiv-cs-ai
23 Jul 2026
Research

Sentence Splitter: Uncovering Latent Factual Structure for Self-Supervised Learning

DGX agent

arXiv:2607.19845v1 Announce Type: cross Abstract: This paper introduces Sentence Splitter, a self-supervised framework built upon a T5-based encoder--decoder architecture for uncovering the latent fac

researcharxiv-cs-ai
23 Jul 2026
Model Releases

SenWorld: A Digital-Twin Simulation for Generating Context-Rich Evaluation Data

DGX agent

arXiv:2607.19949v1 Announce Type: new Abstract: Smartphone personal assistants reason over longitudinal personal data, yet evaluating them requires context-rich evaluation data whose correct answers a

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

Silent Failures in Multimodal Agentic Search:A Diagnostic Taxonomy and Cross-Judge Evaluation

DGX agent

arXiv:2607.19793v1 Announce Type: new Abstract: Multimodal agentic search systems increasingly rely on external tools to answer knowledge-intensive visual questions. However, existing evaluations main

agentsarxiv-cs-ai
23 Jul 2026
Safety

Simulating Eutopia: Revisiting Long-term Fairness with Outcomes, Performativity, and Dynamics

DGX agent

arXiv:2607.19389v1 Announce Type: cross Abstract: As AI-driven Decision Makers (ADMs) influence our socioeconomic reality, their roles in both enhancing efficiency and amplifying the social biases hav

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

DGX agent

arXiv:2607.20145v1 Announce Type: cross Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed trainin

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

SLPO: Scaling Latent Reasoning via a Surrogate Policy

DGX agent

arXiv:2607.19691v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards has become the predominant recipe for eliciting test-time scaling in explicit Chain-of-Thought reasoner

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis

DGX agent

arXiv:2607.20216v1 Announce Type: cross Abstract: Malware analysis demands rapid interpretation of complex detonation reports spanning filesystem, network, and process behaviours. While large language

model-releasesarxiv-cs-ai
23 Jul 2026
← Previous
1…8182838485…448
Next →