AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
8 Jul 2026

Scientific Code Search at Scale: A Multi-Domain Dataset and Benchmark

Model ReleasesDGX agent

arXiv:2607.05443v1 Announce Type: cross Abstract: Scientists increasingly rely on open-source tools to support their research workflows, yet discovering relevant software among over 600 million GitHub

SCOReD: Student-Aware CoT Optimization for Recommendation Distillation

ResearchDGX agent

arXiv:2607.05734v1 Announce Type: cross Abstract: Chain-of-thought (CoT) distillation in the recommendation domain is a necessary precursor to RL training, but raw teacher traces are ill-suited to thi

SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation

SafetyDGX agent

arXiv:2607.05943v1 Announce Type: new Abstract: Training multimodal search agents to perform multi-hop reasoning remains challenging due to a fundamental structural disconnect: existing pipelines cons


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SecureCode: A Production-Grade Multi-Turn Dataset for Training Security-Aware Code Generation Models

AgentsDGX agent

arXiv:2512.18542v3 Announce Type: replace-cross Abstract: AI coding assistants produce vulnerable code in 45% of security-relevant scenarios~ite{veracode2025}, yet no public training dataset teaches b

Segmentation before Answering: Pixel Grounding for MLLM Visual Reasoning

ResearchDGX agent

arXiv:2607.05798v1 Announce Type: cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have evolved from static perception to interleaved visual-language reasoning, often re

Self-Review Reinforcement Learning (SRRL) with Cross-Episode Memory and Policy Distillation

Model ReleasesDGX agent

arXiv:2607.05541v1 Announce Type: cross Abstract: Reinforcement Learning is commonly used to train large language models using environmental feedback. In applied settings, the environment usually prov

Self-Routing: Parameter-Free Expert Routing from Hidden States

Model ReleasesDGX agent

arXiv:2604.00421v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) layers increase model capacity by activating only a small subset of experts per token, and typically rely on a learned rout

Self-Supervised Implicit CEST Reconstruction via Physics-Informed Lorentz Encoding

ResearchDGX agent

arXiv:2607.06132v1 Announce Type: cross Abstract: Multi-Pool Chemical Exchange Saturation Transfer (CEST) MRI provides valuable metabolic information but is clinically limited by long acquisition time

SEVRA-BENCH: Social Engineering of Vulnerabilities in Review Agents

Model ReleasesDGX agent

arXiv:2606.13757v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in automated code-review systems, where their approvals can determine which code is mer

Signed-Graph Recommendation as Structural Consistency Maximization

ResearchDGX agent

arXiv:2607.05952v1 Announce Type: cross Abstract: While signed social recommendation has shown great potential by modeling both trust and distrust relations, its effectiveness is often hindered by str

Sparse but Wrong: Incorrect L0 Leads to Incorrect Features in Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2508.16560v4 Announce Type: replace-cross Abstract: Sparse Autoencoders (SAEs) extract features from LLM internal activations, meant to correspond to interpretable concepts. A core SAE training

SpatialFly: Implicit 3D Prior-Guided Visual Reparameterization for Continuous UAV Vision-and-Language Navigation

SafetyDGX agent

arXiv:2603.21046v2 Announce Type: replace-cross Abstract: UAVs play an important role in applications such as autonomous exploration, disaster response, and infrastructure inspection. However, UAV VLN

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows

Model ReleasesDGX agent

arXiv:2607.06229v1 Announce Type: cross Abstract: Major cloud data platforms now expose large language model capabilities as native SQL functions, enabling analysts to perform classification, filterin

StateFuse: Deterministic Conflict-Preserving Memory for Multi-Agent Systems

AgentsDGX agent

arXiv:2607.05844v1 Announce Type: new Abstract: Agent systems accumulate conflicting observations across branches, retries, and replicas, yet many practical memory layers still collapse disagreement b

Static Metrics Are Insufficient: Predicting Java Method Energy Usage with Execution Time

ResearchDGX agent

arXiv:2607.06124v1 Announce Type: cross Abstract: The increasing energy demand of software systems is raising concerns about their environmental impact and associated costs. Reasoning on energy usage

Statistical Adversaries: Natural Backdoor-like Features in Vision Datasets

SafetyDGX agent

arXiv:2607.05516v1 Announce Type: cross Abstract: Model-specific adversarial attacks have been extensively studied. We study a different failure mode: naturally occurring statistical signals in vision

StepShield: When, Not Whether to Intervene on Rogue Agents

Model ReleasesDGX agent

arXiv:2601.22136v2 Announce Type: replace-cross Abstract: Agent safety benchmarks measure whether a monitor detects harm, not when. Yet timing is the difference between intervention and autopsy. We in

Synthetic Consumer Insight Generation with Large Language Models

TutorialsDGX agent

arXiv:2607.05761v1 Announce Type: new Abstract: Modern data-driven marketing relies on large amounts of consumer data, yet collecting such data can be costly, time-consuming, and difficult to scale. T

Tangent classes of matroids and wonderful compactifications

AgentsDGX agent

arXiv:2607.05835v1 Announce Type: cross Abstract: For every loopless matroid M and every Feichtner--Yuzvinsky building set G containing the top flat, we construct an integral tangent class T_{M,G}^{Z}

Task Decomposition-Guided Reranking for Adaptive Agent Skill Retrieval

AgentsDGX agent

arXiv:2607.06283v1 Announce Type: new Abstract: Skill usage can significantly enhance the ability of modern agent systems to complete complex tasks. However, the growing scale of skill libraries makes

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

Model ReleasesDGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

The GenAI Skill Bypass: Mapping Divergent Pathways of University Students and Staff AI Literacy

ApplicationsDGX agent

arXiv:2607.05411v1 Announce Type: cross Abstract: Higher education institutions are increasingly expected to ensure that both students and staff develop Generative AI (GenAI) literacies. In response,

The Granularity Paradox: How Temporal Disaggregation Inflates In-Sample Fit and Compounds Out-of-Sample Error

Model ReleasesDGX agent

arXiv:2607.05450v1 Announce Type: cross Abstract: This paper explores the 'Granularity Paradox' in time-series forecasting, wherein finer temporal disaggregation (e.g., Monthly to Weekly/Daily) improv

The Jagged Global Economy: Frontier AI Unevenly Exposes National Economies

SafetyDGX agent

arXiv:2607.05404v1 Announce Type: cross Abstract: Frontier AI's labor-market effects matter to workers, firms, and policymakers, but current evidence generally comes from a handful of high-income econ

The Large Cancer Assistant (LCA): A Model-Agnostic Orchestration Framework for Scalable Clinical Decision Support in Oncology

SafetyDGX agent

arXiv:2607.06531v1 Announce Type: new Abstract: - Objective: Multimodal deep learning models in oncology are currently limited by monolithic designs that rigidly couple data ingestion, clinical routin

The relationship between reasoning and performance in large language models--o3 (mini) thinks harder, not longer

Model ReleasesDGX agent

arXiv:2502.15631v2 Announce Type: replace-cross Abstract: Large language models have demonstrated remarkable progress in mathematical reasoning, leveraging chain-of-thought and reinforcement learning.

The yes-no bias of large language models reflects answer order and wording, not shifts in moral judgment

Model ReleasesDGX agent

arXiv:2607.05552v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly issue judgments read as binary verdicts, and a growing literature reports such judgments shifting under logi

Think Before You Grid-Search: Floor-First Triage for LLM Serving

Model ReleasesDGX agent

arXiv:2607.05876v1 Announce Type: cross Abstract: LLM serving optimization typically benchmarks many configurations and reaches for heavy profilers when latency targets are missed. We argue for the re

TILDE: TILt-based Distributional Erasure for Concept Unlearning

SafetyDGX agent

arXiv:2607.06432v1 Announce Type: cross Abstract: Concept unlearning in text-to-image diffusion models is critical for safe and practical deployment: with rising privacy concerns, copyright disputes,

To Retain or to Adapt? Generalizing Continual Learning

SafetyDGX agent

arXiv:2607.05609v1 Announce Type: cross Abstract: The Continual Learning (CL) literature has long been driven by the goal of mitigating catastrophic forgetting. This objective rests on a pervasive, of

Token-Based Dual-view Fusion and Adaptation of Large Vision Models for Breast Cancer Classification

ResearchDGX agent

arXiv:2607.06309v1 Announce Type: cross Abstract: Accurate breast cancer classification from mammography requires effective integration of complementary information from craniocaudal (CC) and mediolat

TopoBrick: Agentic Topology Sampling of Exogenous Variables for Zero-Shot Building IoT Forecasting

AgentsDGX agent

arXiv:2607.06349v1 Announce Type: new Abstract: Building sensors are embedded in physical topology, spatial hierarchy, and operational context, yet existing forecasters often treat them as isolated ti

Toward AI standardization: A triadic human-ai collaboration framework for multi-level autonomous mobility

AgentsDGX agent

arXiv:2504.19120v2 Announce Type: replace-cross Abstract: The goal of the current study is to introduce a triadic human-AI collaboration framework that could be applied in transportation systems such

Transformers converge to invariant algorithmic cores

Model ReleasesDGX agent

arXiv:2602.22600v2 Announce Type: replace-cross Abstract: Training selects for behavior, not circuitry: many weight configurations can implement the same function. Studying any single trained neural n

TriA Pipeline: A Large-Scale Automatic Audio Annotation Pipeline For Audio Classification In Specific Scenarios

ResearchDGX agent

arXiv:2607.06179v1 Announce Type: cross Abstract: There are some datasets of varying scales for audio classification (AC) applied to different tasks. However, annotated data is limited for most scenar

Trust-free Personalized Decentralized Learning

Local AiDGX agent

arXiv:2410.11378v3 Announce Type: replace-cross Abstract: Personalized collaborative learning in federated settings faces a critical trade-off between customization and participant trust. Existing app

TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training

SafetyDGX agent

arXiv:2607.05804v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student policy by matching a stronger teacher on the student's own trajectories, offering a promising framework fo

UBEP: Re-architecting Expert Parallelism Communication Library for Production Superpods

HardwareDGX agent

arXiv:2607.06202v1 Announce Type: cross Abstract: The deployment of Mixture-of-Experts (MoE) models on production high-bandwidth superpods, such as NVIDIA's NVL72/576 and Huawei's CloudMatrix384, intr

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

Model ReleasesDGX agent

arXiv:2607.06306v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex pro

Uncovering Latent Depression Severity for Binary Depression Detection via Advantage-weighting Ranking

ResearchDGX agent

arXiv:2607.05901v1 Announce Type: new Abstract: Automatic depression detection using audio-visual data faces significant challenges, particularly in disentangling overlapping feature distributions and

Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An Approval-View Fidelity Gap Across Three Independent Server Implementations

AgentsDGX agent

arXiv:2607.05744v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is the dominant way coding agents discover and invoke external tools. A server advertises each tool through a tools/l

Universal Algorithm-Implicit Learning

ResearchDGX agent

arXiv:2602.14761v2 Announce Type: replace-cross Abstract: Current meta-learning methods are constrained to narrow task distributions with fixed feature and label spaces, limiting applicability. Moreov

Unsupervised Anomaly Detection of Information Operations Users via Behavioral and Language Patterns

ApplicationsDGX agent

arXiv:2607.05855v1 Announce Type: cross Abstract: Information Operations on social media networks have been identified as a significant threat to democracy and modern society, but they are challenging

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations

AgentsDGX agent

arXiv:2512.19458v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly embedded in agentic frameworks for scientific discovery. First-principles materials computation impose

VendorBench-100: A Unified Cross-Paradigm Benchmark for Deepfake Image Detection

Model ReleasesDGX agent

arXiv:2607.06254v1 Announce Type: cross Abstract: Deepfake image detection is currently served by three fundamentally different paradigms: commercial APIs, zero-shot vision-language models (LLMs), and

VisTCP: A Visualization Framework to Construct Knowledge-Graph-Based Representation for Traditional Chinese Painting

ApplicationsDGX agent

arXiv:2607.05841v1 Announce Type: cross Abstract: Structured representation can characterize semantic objects and relationships in images. It provides a possible effective way for the semantic underst

Volumetric Directional Diffusion: Anchoring Uncertainty Quantification in Anatomical Consensus for Ambiguous Medical Image Segmentation

SafetyDGX agent

arXiv:2603.04024v2 Announce Type: replace-cross Abstract: Ambiguous 3D medical image segmentation often involves boundaries where different expert delineations are non-identical yet clinically plausib

What Counts as Real? Speech Restoration and Voice Quality Conversion Pose New Challenges to Deepfake Detection

ResearchDGX agent

arXiv:2603.14033v2 Announce Type: replace-cross Abstract: Audio anti-spoofing systems are typically trained to assign one authenticity label to an entire speech utterance. This formulation becomes und

What Do AI Agents Actually Change? An Empirical Taxonomy of Mutation Patterns in Performance-Improving Pull Requests

AgentsDGX agent

arXiv:2607.05666v1 Announce Type: cross Abstract: AI coding agents are black boxes: we cannot inspect how they generate code, but we can inspect what they change. This distinction matters for search-b

What Images Cannot Say: Language-Guided Olfactory Representation Learning

ResearchDGX agent

arXiv:2607.06402v1 Announce Type: cross Abstract: Images tell us what a scene looks like, but rarely what it would feel like to be there. While recent datasets pair visual scenes with electronic-nose

When AI Classifies: What Counts as Public Administration?

ResearchDGX agent

arXiv:2607.05420v1 Announce Type: cross Abstract: This study examines how alternative systems of scholarly representation identify and characterize broad public administration (PA) and artificial inte

When Assisting One Disempowers Another

AgentsDGX agent

arXiv:2511.04177v2 Announce Type: replace Abstract: Personal AI agents are increasingly deployed in shared environments, where their actions affect not just the primary user they are assisting, but by

When do prophets profit in prediction markets?

ResearchDGX agent

arXiv:2607.06166v1 Announce Type: new Abstract: Prediction markets aggregate dispersed beliefs into prices that act as probabilistic forecasts of uncertain events. Classical theory establishes a clean

When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents

SafetyDGX agent

arXiv:2606.20023v2 Announce Type: replace-cross Abstract: As LLM agents increasingly select tools autonomously, their choices among tools with different privileges become safety-relevant. However, pri

When Should LLMs Search? Counterfactual Supervision for Search Routing

Model ReleasesDGX agent

arXiv:2607.05752v1 Announce Type: cross Abstract: Search-augmented language models can use external evidence to compensate for limitations in parametric knowledge, but search is not uniformly benefici

Whose fairness? Structural concentration in AI bias research

SafetyDGX agent

arXiv:2607.05574v1 Announce Type: cross Abstract: Artificial intelligence increasingly mediates consequential decisions in healthcare, law, and public services, and the field has responded with an ext

Why does AI unlock new possibilities in STEM education? A Bibliometric Analysis of Trends and Future Agenda

ApplicationsDGX agent

arXiv:2607.05412v1 Announce Type: cross Abstract: STEM education faces challenges in personalization and interdisciplinary integration. AI technology has brought new possibilities, but the mechanisms

X-FEMR: A Token-level Explainable Approach for Electronic Health Records Foundation Models using Transformer-based Models

SafetyDGX agent

arXiv:2607.06163v1 Announce Type: cross Abstract: Foundation Models for Electronic Health Records (FEMRs) are pretrained on large-scale structured patient data, enabling them to convert longitudinal p

x-Prediction Is All You Need:Training-Free Accelerated Generation via Endpoint Decodability

Model ReleasesDGX agent

arXiv:2607.06114v1 Announce Type: cross Abstract: Diffusion and flow matching models generate high-quality samples, but their ODE samplers often need tens to hundreds of neural function evaluations (N

7 Jul 2026

A Bayesian Framework for Evaluating Scenario Compatibility in Generative Population Synthesis

SafetyDGX agent

arXiv:2607.03190v1 Announce Type: cross Abstract: Scenario-based transportation analysis specifies future assumptions through aggregate population targets, whereas generative population synthesis mode

← Previous
1…8283848586…358
Next →