AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
28 Apr 2026

Progressive Approximation in Deep Residual Networks: Theory and Validation

Model ReleasesDGX agent

arXiv:2604.24154v1 Announce Type: cross Abstract: The Universal Approximation Theorem (UAT) guarantees universal function approximation but does not explain how residual models distribute approximatio

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

SafetyDGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

PushupBench: Your VLM is not good at counting pushups

ResearchDGX agent

arXiv:2604.23407v1 Announce Type: cross Abstract: Large vision-language models (VLMs) can recognize extit{what} happens in video but fail to count extit{how many} times. We introduce extbf{PushupBench


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

QED: An Open-Source Multi-Agent System for Generating Mathematical Proofs on Open Problems

Model ReleasesDGX agent

arXiv:2604.24021v1 Announce Type: new Abstract: We explore a central question in AI for mathematics: can AI systems produce original, nontrivial proofs for open research problems? Despite strong bench

QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering

Model ReleasesDGX agent

arXiv:2604.24052v1 Announce Type: cross Abstract: Video-to-text summarization remains underexplored in terms of comprehensive evaluation methods. Traditional n-gram overlap-based metrics and recent la

Quantifying and Improving the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data

ApplicationsDGX agent

arXiv:2503.05587v3 Announce Type: replace-cross Abstract: Robustness has become a critical attribute for the deployment of RAG systems in real-world applications. Existing research focuses on robustne

Quantifying and Mitigating Self-Preference Bias of LLM Judges

SafetyDGX agent

arXiv:2604.22891v1 Announce Type: cross Abstract: LLM-as-a-Judge has become a dominant approach in automated evaluation systems, playing critical roles in model alignment, leaderboard construction, qu

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

SafetyDGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

Quantum Kernel Advantage over Classical Collapse in Medical Foundation Model Embeddings

ResearchDGX agent

arXiv:2604.24597v1 Announce Type: cross Abstract: We provide evidence of quantum kernel advantage under noiseless simulation in binary insurance classification on MIMIC-CXR chest radiographs using qua

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

Model ReleasesDGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

Quasi-Quadratic Gradient: A New Direction for Accelerating the BFGS Method in Quasi-Newton Optimization

ResearchDGX agent

arXiv:2604.23922v1 Announce Type: cross Abstract: In this paper, we introduce the Quasi-Quadratic Gradient (QQG), a novel search direction designed to accelerate the BFGS method within the quasi-Newto

Query2Diagram: Answering Developer Queries with UML Diagrams

ResearchDGX agent

arXiv:2604.23816v1 Announce Type: cross Abstract: Software documentation frequently becomes outdated or fails to exist entirely, yet developers need focused views of their codebase to understand compl

Question-Adaptive Graph Learning for Multi-hop Retrieval Augmented Generation

TutorialsDGX agent

arXiv:2510.11541v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has demonstrated its ability to enhance Large Language Models (LLMs) by integrating external knowledge so

RADIANT-LLM: an Agentic Retrieval Augmented Generation Framework for Reliable Decision Support in Safety-Critical Nuclear Engineering

Local AiDGX agent

arXiv:2604.22755v1 Announce Type: cross Abstract: Reliable decision support in nuclear engineering requires traceable, domain-grounded knowledge retrieval, yet safety and risk analysis workflows remai

RAS: a Reliability Oriented Metric for Automatic Speech Recognition

Model ReleasesDGX agent

arXiv:2604.24278v1 Announce Type: cross Abstract: Automatic speech recognition systems often produce confident yet incorrect transcriptions under noisy or ambiguous conditions, which can be misleading

RAT: RunAnyThing via Fully Automated Environment Configuration

Model ReleasesDGX agent

arXiv:2604.23190v1 Announce Type: cross Abstract: Automating repository-level software engineering tasks is a foundational challenge for autonomous code agents, largely due to the difficulty of config

RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing

Model ReleasesDGX agent

arXiv:2604.23644v1 Announce Type: cross Abstract: Intelligent document processing pipelines extract structured entities (tables, images, and text) from documents for use in downstream systems such as

RCSB PDB AI Help Desk: retrieval-augmented generation for protein structure deposition support

Model ReleasesDGX agent

arXiv:2604.22800v1 Announce Type: cross Abstract: Motivation: Structural Biologists have contributed more than 245,000 experimentally determined three-dimensional structures of biological macromolecul

RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?

Model ReleasesDGX agent

arXiv:2602.07096v2 Announce Type: replace-cross Abstract: Reliable financial reasoning requires knowing not only how to answer, but also when an answer cannot be justified. In real financial practice,

Reasonably reasoning AI agents can avoid game-theoretic failures in zero-shot, provably

Model ReleasesDGX agent

arXiv:2603.18563v2 Announce Type: replace Abstract: As autonomous AI agents increasingly mediate online platform markets, a fundamental question emerges: do these markets generate stable strategic out

Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization

Model ReleasesDGX agent

arXiv:2408.00923v2 Announce Type: replace-cross Abstract: This paper explores a novel paradigm in low-bit (i.e. 4-bits or lower) quantization, differing from existing state-of-the-art methods, by fram

Reconstructive Authority Model: Runtime Execution Validity Under Partial Observability

AgentsDGX agent

arXiv:2604.22898v1 Announce Type: cross Abstract: Autonomous systems increasingly operate under partial observability where execution-relevant state is never fully accessible. Existing governance mech

RedParrot: Accelerating NL-to-DSL for Business Analytics via Query Semantic Caching

ApplicationsDGX agent

arXiv:2604.22758v1 Announce Type: cross Abstract: Recently, at Xiaohongshu, the rapid expansion of e-commerce and advertising demands real-time business analytics with high accuracy and low latency. T

RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation

Model ReleasesDGX agent

arXiv:2604.24218v1 Announce Type: cross Abstract: As the complexity of System-on-Chip (SoC) designs grows, the shift-left paradigm necessitates the rapid development of high-fidelity reference models

ReFinE: Streamlining UI Mockup Iteration with Research Findings

TutorialsDGX agent

arXiv:2604.04353v2 Announce Type: replace-cross Abstract: Although HCI research papers offer valuable design insights, designers often struggle to apply them in design workflows due to difficulties in

Reflective Flow Sampling Enhancement

SafetyDGX agent

arXiv:2603.06165v2 Announce Type: replace-cross Abstract: The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models tra

Reheat Nachos for Dinner? Evaluating AI Support for Cross-Cultural Communication of Neologisms

TutorialsDGX agent

arXiv:2604.23842v1 Announce Type: cross Abstract: Neologisms and emerging slang are central to daily conversation, yet challenging for non-native speakers (NNS) to interpret and use appropriately in c

Reinforcement Learning with Backtracking Feedback

Model ReleasesDGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

Reliable Microservice Tail Latency Prediction via Decoupled Dual-Stream Learning and Gradient Modulation

Local AiDGX agent

arXiv:2508.01635v2 Announce Type: replace-cross Abstract: Microservice architectures enable scalable cloud-native applications; however, the distributed nature of these systems complicates the mainten

Representation Homogeneity and Systemic Instability in AI-Dominated Financial Markets: A Structural Approach

AgentsDGX agent

arXiv:2604.22818v1 Announce Type: cross Abstract: This paper investigates how similarity in the informational representation of market states among Artificial Intelligence (AI) trading agents can gene

Representational Curvature Modulates Behavioral Uncertainty in Large Language Models

TutorialsDGX agent

arXiv:2604.23985v1 Announce Type: new Abstract: In autoregressive large language models (LLMs), temporal straightening offers an account of how the next-token prediction objective shapes representatio

ResAF-Net: An Anchor-Free Attention-Based Network for Tree Detection and Agricultural Mapping in Palestine

Model ReleasesDGX agent

arXiv:2604.23653v1 Announce Type: cross Abstract: Reliable agricultural data is essential for food security, land-use planning, and economic resilience, yet in Palestine, such data remains difficult t

Resolution scaling governs DINOv3 transfer performance in chest radiograph classification

Model ReleasesDGX agent

arXiv:2510.07191v3 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has improved visual representation learning, but its value in chest radiography remains uncertain. DINOv3 exten

Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs

Model ReleasesDGX agent

arXiv:2509.25414v2 Announce Type: replace-cross Abstract: Large language models are often adapted using parameter-efficient techniques such as Low-Rank Adaptation (LoRA), formulated as y = W_0x + BAx,

Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes

SafetyDGX agent

arXiv:2603.25562v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) is increasingly used in LLM post-training because it can leverage a teacher model to provide dense supervision on

Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis

SafetyDGX agent

arXiv:2604.24198v1 Announce Type: cross Abstract: Process Reward Models (PRMs) have achieved remarkable success in augmenting the reasoning capabilities of Large Language Models (LLMs) within static d

Right-to-Act: A Pre-Execution Non-Compensatory Decision Protocol for AI Systems

SafetyDGX agent

arXiv:2604.24153v1 Announce Type: new Abstract: Current AI systems increasingly operate in contexts where their outputs directly trigger real-world actions. Most existing approaches to AI safety, risk

Risk-Aware Robust Learning: Reducing Clinical Risk under Label Noise in Medical Image Classification

SafetyDGX agent

arXiv:2604.23875v1 Announce Type: cross Abstract: Noisy labels are a pervasive challenge in medical image classification, where annotation errors arise from inter-observer variability and diagnostic a

RL-ASL: A Dynamic Listening Optimization for TSCH Networks Using Reinforcement Learning

ResearchDGX agent

arXiv:2604.07533v2 Announce Type: replace-cross Abstract: Time Slotted Channel Hopping (TSCH) is a widely adopted Media Access Control (MAC) protocol within the IEEE 802.15.4e standard, designed to pr

RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents

SafetyDGX agent

arXiv:2604.22888v1 Announce Type: cross Abstract: Agent skills introduce a new and more severe form of indirect injection for LLM agents: unlike traditional indirect prompt injection, attackers can hi

S2G-RAG: Structured Sufficiency and Gap Judging for Iterative Retrieval-Augmented QA

ResearchDGX agent

arXiv:2604.23783v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds language models in external evidence, but multi-hop question answering remains difficult because iterativ

S^2IT: Stepwise Syntax Integration Tuning for Large Language Models in Aspect Sentiment Quad Prediction

Local AiDGX agent

arXiv:2604.23296v1 Announce Type: cross Abstract: Aspect Sentiment Quad Prediction (ASQP) has seen significant advancements, largely driven by the powerful semantic understanding and generative capabi

Scalable Agentic Reasoning for Designing Biologics Targeting Intrinsically Disordered Proteins

Model ReleasesDGX agent

arXiv:2512.15930v2 Announce Type: replace-cross Abstract: Intrinsically disordered proteins (IDPs) represent crucial therapeutic targets due to their significant role in disease -- approximately 80% o

Scalable Explainability-as-a-Service (XaaS) for Edge AI Systems

Local AiDGX agent

arXiv:2602.04120v2 Announce Type: replace-cross Abstract: Though Explainable AI (XAI) has made significant advancements, its inclusion in edge and IoT systems is typically ad-hoc and inefficient. Most

Scalable Hyperparameter-Divergent Ensemble Training with Automatic Learning Rate Exploration for Large Models

HardwareDGX agent

arXiv:2604.24708v1 Announce Type: cross Abstract: Training large neural networks with data-parallel stochastic gradient descent allocates N GPU replicas to compute effectively identical updates -- a p

Scalable LLM-based Coding of Dialogue in Healthcare Simulation: Balancing Coding Performance, Processing Time, and Environmental Impact

ApplicationsDGX agent

arXiv:2604.23255v1 Announce Type: cross Abstract: Research shows that dialogue, the interactive process through which participants articulate their thinking, plays a central role in constructing share

Scalable Production Scheduling: Linear Complexity via Unified Homogeneous Graphs

SafetyDGX agent

arXiv:2604.23841v1 Announce Type: cross Abstract: Efficiently solving the Job Shop Scheduling Problem in real-world industrial applications requires policies that are both computationally lean and top

Scaling Coding Agents via Atomic Skills

ResearchDGX agent

arXiv:2604.05013v2 Announce Type: replace-cross Abstract: Current LLM coding agents are predominantly trained on composite benchmarks (e.g., bug fixing), which often leads to task-specific overfitting

Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns

Model ReleasesDGX agent

arXiv:2604.23150v1 Announce Type: cross Abstract: Most recent state-of-the-art (SOTA) large language models (LLMs) use Mixture-of-Experts (MoE) architectures to scale model capacity without proportion

Scaling Properties of Continuous Diffusion Spoken Language Models

Model ReleasesDGX agent

arXiv:2604.24416v1 Announce Type: cross Abstract: Speech-only spoken language models (SLMs) lag behind text and text-speech models in performance, with recent discrete autoregressive (AR) SLMs indicat

Scheduling Your LLM Reinforcement Learning with Reasoning Trees

SafetyDGX agent

arXiv:2510.24832v2 Announce Type: replace Abstract: Using Reinforcement Learning with Verifiable Rewards (RLVR) to optimize Large Language Models (LLMs) can be conceptualized as progressively editing

Scheming Ability in LLM-to-LLM Strategic Interactions

Model ReleasesDGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

Scoring, Reasoning, and Selecting the Best! Ensembling Large Language Models via a Peer-Review Process

ResearchDGX agent

arXiv:2512.23213v3 Announce Type: replace-cross Abstract: We propose LLM-PeerReview, an unsupervised LLM Ensemble method that selects the most ideal response from multiple LLM-generated candidates for

ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules

Model ReleasesDGX agent

arXiv:2603.29928v2 Announce Type: replace Abstract: Tabular foundation models such as TabPFN and TabICL already produce full predictive distributions, yet prevailing regression benchmarks evaluate the

SCRIBE: Structured Mid-Level Supervision for Tool-Using Language Models

AgentsDGX agent

arXiv:2601.03555v2 Announce Type: replace Abstract: Training reliable tool-augmented agents remains a significant challenge, largely due to the difficulty of credit assignment in multi-step reasoning.

SeaEvo: Advancing Algorithm Discovery with Strategy Space Evolution

AgentsDGX agent

arXiv:2604.24372v1 Announce Type: cross Abstract: LLM-guided evolutionary search has emerged as a promising paradigm for automated algorithm discovery, yet most systems track search progress primarily

Security Considerations for Multi-agent Systems

SafetyDGX agent

arXiv:2603.09002v2 Announce Type: replace-cross Abstract: Multi-agent artificial intelligence systems or MAS are systems of autonomous agents that exercise delegated tool authority, share persistent m

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

Model ReleasesDGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

See No Evil: Semantic Context-Aware Privacy Risk Detection for AR

ApplicationsDGX agent

arXiv:2604.22805v1 Announce Type: cross Abstract: Augmented reality (AR) systems pose unique privacy risks due to their continuous capture of visual data. Existing AR privacy frameworks lack semantic

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

Model ReleasesDGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

← Previous
1…301302303304305…354
Next →