AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
28 Jul 2026

DataOrchestra: Learning to Orchestrate Per-Example Curation of Pretraining Data

ResearchDGX agent

arXiv:2607.24717v1 Announce Type: cross Abstract: Pretraining data processing is critical to the downstream performance of Large Language Models (LLMs). However, many existing approaches define a fixe

Decentralized Causal Discovery using Judo Calculus

ApplicationsDGX agent

arXiv:2510.23942v2 Announce Type: replace Abstract: We describe a theory and implementation of an intuitionistic decentralized framework for causal discovery using judo calculus, which is formally def

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

Model ReleasesDGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DecoupleMix: Decoupled Ratio Search and Convex Allocation for Scalable VLM Data Recipes

ResearchDGX agent

arXiv:2607.24516v1 Announce Type: cross Abstract: While data curation for Vision Language Models (VLMs) is increasingly active, public practice for constructing pretraining mixtures remains largely he

DeepFaith: Evidence-Grounded LLMs for Faithful Incident Reporting in Multi-Stage APT Defense

AgentsDGX agent

arXiv:2607.24348v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) are difficult to detect and interpret due to their multi-stage and stealthy nature. While recent autonomous defense

DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs

Model ReleasesDGX agent

arXiv:2607.22555v1 Announce Type: new Abstract: Medical diagnosis is a multi-stage process: extract facts, consult knowledge, generate a differential analysis, and select the best diagnosis with expla

DeepLook: Deeper Thinking with Lookahead

Model ReleasesDGX agent

arXiv:2607.22602v1 Announce Type: new Abstract: Inference-time scaling has emerged as a powerful paradigm for improving large language model reasoning, often delivering larger gains on difficult reaso

Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis

AgentsDGX agent

arXiv:2607.23524v1 Announce Type: new Abstract: Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This couple

Denial of Deadline: Network-Driven Accuracy Collapse in Distributed Inference Pipelines

AgentsDGX agent

arXiv:2607.24692v1 Announce Type: cross Abstract: Inference systems increasingly combine a fast path that returns predictions within the application's latency deadline together with a higher-accuracy

Design Theater: A Benchmark for Generative UI

Model ReleasesDGX agent

arXiv:2607.22928v1 Announce Type: new Abstract: Generative UI tools promise to democratize UI design by turning natural language descriptions into complete interfaces. Alongside the interface, these t

Designing Service Systems from Textual Evidence

SafetyDGX agent

arXiv:2603.10400v2 Announce Type: replace-cross Abstract: Designing service systems requires selecting among alternative configurations -- choosing the best chatbot variant, the optimal routing policy

DICA: Dual-Indicator Guided Contrastive Alignment in Multimodal Large Language Models

SafetyDGX agent

arXiv:2607.23944v1 Announce Type: new Abstract: Human visual reasoning typically follows a coarse-to-fine attention process, starting from global scene understanding and gradually focusing on question

Differencing the Diffusion Trajectory toward Uncertain Components for Time Series Forecasting

ResearchDGX agent

arXiv:2607.22599v1 Announce Type: new Abstract: Diffusion models have become a widely used framework for probabilistic time series forecasting, modeling the distribution of future values given an obse

Directional Influence Function: Estimating Training Data Influence in Constrained Learning

SafetyDGX agent

arXiv:2607.23388v1 Announce Type: cross Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustnes

Disentangling Multi-View Scanning in Mamba for Network Traffic Anomaly Detection

ResearchDGX agent

arXiv:2607.22829v1 Announce Type: new Abstract: Network Traffic Anomaly Detection (NTAD) is a critical task in cybersecurity, yet timely and accurate anomaly detection remains challenging. Mamba has e

Disentangling Semantic Attention from Structural Bias in the Attention Manifold

SafetyDGX agent

arXiv:2607.24017v1 Announce Type: cross Abstract: The empirical success of attention mechanism in Multimodal Large Language Models (MLLMs) often obscures its inherent, subtle flaws. Specifically, MLLM

Do Coverage and Mutation Scores of LLM-Generated Test Suites Correlate with Their Effectiveness? (Replicability Study)

TutorialsDGX agent

arXiv:2607.22880v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have driven growing interest in using LLMs to automate test generation. Prior work commonly evaluates

Do Diagrams Help Large Language Models Reason? Evidence from Syllogistic Reasoning

Model ReleasesDGX agent

arXiv:2607.23513v1 Announce Type: cross Abstract: Diagrams are widely used to support logical reasoning, and prior studies suggest that representations such as Euler diagrams can improve human reasoni

Do Language Models Converge to Themselves? Recursive Self-Refinement as Textual Relaxation

Model ReleasesDGX agent

arXiv:2607.22653v1 Announce Type: new Abstract: Large language models are increasingly used in recursive refinement workflows, where an initial draft is repeatedly revised by the same model. Despite t

Do LLMs Know Their Vulnerable Scenarios?

Model ReleasesDGX agent

arXiv:2607.23496v1 Announce Type: new Abstract: Safety-aligned large language models are trained to refuse harmful requests, yet embedding the same requests in particular scenarios can bypass their sa

Do Small Models Use the Law You Give Them? Context-Injected Fine-Tuning for Legal QA in Bangladesh

ApplicationsDGX agent

arXiv:2607.23446v1 Announce Type: cross Abstract: A small language model can receive the governing statutory provision and still answer incorrectly. We test whether fine-tuning on examples containing

Do Visual Features Improve Other-Initiated Repair Detection? A Dyadic Multimodal Approach

ResearchDGX agent

arXiv:2607.23845v1 Announce Type: new Abstract: Other-initiated Self-repair, or in short Other-initiated Repair (OIR), is an essential mechanism in conversational interaction, whereby a recipient sign

DocHRL: A Hierarchical Reinforcement Learning Framework for Cost-Optimised Document Classification

Model ReleasesDGX agent

arXiv:2607.22644v1 Announce Type: new Abstract: Real-world document classification pipelines typically apply the same sequence of models to every incoming document, regardless of its complexity or typ

DomainPilot: Domain-Level Loss-Guided Two-Stage Data Mixture Optimization for Efficient Language Model Fine-Tuning

ResearchDGX agent

arXiv:2607.22769v1 Announce Type: cross Abstract: The training efficacy of large language models (LLMs) is fundamentally constrained by the quality and composition of training data. Existing dynamic d

DOSA: A Tree-Guided, Self-Regressive Framework for Long Document Structure Analysis

Model ReleasesDGX agent

arXiv:2607.22679v1 Announce Type: new Abstract: In visually-rich documents, information is encoded not only in individual page objects such as tables, headers, and text blocks, but also in the structu

DraftExpert: Expansion-Aware Self-Speculative Decoding for End-Device MoE Inference

Model ReleasesDGX agent

arXiv:2607.24434v1 Announce Type: cross Abstract: Large Mixture-of-Experts (MoE) language models are attractive for end-device deployment because only a small subset of experts is active per token, bu

DreamCAD: Scaling Multi-modal CAD Generation using Differentiable Parametric Surfaces

Model ReleasesDGX agent

arXiv:2603.05607v2 Announce Type: replace-cross Abstract: Computer-Aided Design (CAD) relies on structured and editable geometric representations, yet existing generative methods are constrained by sm

DSCH-Loss: A Dynamic Semantic Channel Objective for Deep Semantic Hashing

ResearchDGX agent

arXiv:2607.24567v1 Announce Type: new Abstract: Semantic hashing methods for generating short binary hash codes that allow efficient approximate nearest neighbor search in high-dimensional data spaces

DSTFView: Multi-View Cloud-Edge Workload Forecasting with Dual-Input Spatio-Temporal-Frequency Modeling

ResearchDGX agent

arXiv:2607.22565v1 Announce Type: new Abstract: With the widespread deployment of edge-side AI inference, edge platforms are increasingly required to support latency-sensitive, highly concurrent, and

DualityCert: Verifier-Gated Language-Model Repair of Broken Duality Claims in Quantum Field Theory

Model ReleasesDGX agent

arXiv:2607.23614v1 Announce Type: cross Abstract: We present DualityCert, a symbolic verifier for candidate Seiberg-duality claims in four-dimensional N=1 quiver gauge theories. The verifier evaluates

DuoAD: Leveraging [CLS] Dual Characteristics for Training-Free Few-Shot Anomaly Detection

Model ReleasesDGX agent

arXiv:2607.23924v1 Announce Type: cross Abstract: Vision foundation models have enabled strong training-free anomaly detection (AD). However, most existing approaches rely primarily on independent loc

DynaResize: Runtime GPU Reallocation for Disaggregated LLM Post-Training

HardwareDGX agent

arXiv:2607.22614v1 Announce Type: new Abstract: RL-based LLM post-training increasingly disaggregates Rollout and Training across separate GPU resources, but static GPU partitioning suffers from sever

E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios

Model ReleasesDGX agent

arXiv:2607.23722v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as agents that interact with stateful environments over multiple steps: gathering hidden informat

Earnings25: A Comprehensive 500-Hour Speech Benchmark for Finance

Model ReleasesDGX agent

arXiv:2607.23813v1 Announce Type: cross Abstract: We introduce Earnings25, a finance-domain benchmark for evaluating automatic speech recognition (ASR) on English-language earnings calls under realist

EchoBridge: Long-Tail-Aware ECG-Echocardiography Text Alignment for Echocardiography-Derived Cardiac Findings

SafetyDGX agent

arXiv:2607.24553v1 Announce Type: cross Abstract: Standardized echocardiography conclusions provide meaningful supervision for learning ECG representations of echocardiography-derived cardiac findings

EEGForceFusion: Joint Tokenised-Continuous Representation Learning for Subject-Independent Grasp Force Decoding

ResearchDGX agent

arXiv:2607.24126v1 Announce Type: cross Abstract: Brain-machine interfaces provide a link between neural activity and external devices, enabling restoration of motor function and advancing human-machi

Efficiency Matters in Autonomous Research

SafetyDGX agent

arXiv:2607.24647v1 Announce Type: new Abstract: AI-driven autonomous research (AR) systems are becoming increasingly effective across a broad range of tasks. Their performance, however, is still evalu

Efficient LLM-Generated Shuttling Compilers for Complex Trapped-Ion Architectures

Model ReleasesDGX agent

arXiv:2607.24714v1 Announce Type: cross Abstract: Trapped-ion quantum computers rely on shuttling compilers, which cast an input algorithm into a sequence of ion-qubit movements within a given archite

EgoPlay: Event-Triggered Video Editing for Egocentric Streams

Model ReleasesDGX agent

arXiv:2607.24560v1 Announce Type: cross Abstract: We introduce EgoPlay, an event-triggered video-to-video editor for egocentric streams, obtained by fine-tuning a pretrained V2V diffusion transformer

Embodied GPT-5.1: Evidence of a World Model?

Model ReleasesDGX agent

arXiv:2607.23899v1 Announce Type: cross Abstract: This exploratory study examines whether a large multimodal language model, GPT-5.1, can serve as the high-level controller of a physical mobile robot

Epistemic Norms for AI Safety and Alignment Research

SafetyDGX agent

arXiv:2607.24243v1 Announce Type: new Abstract: Mainstream AI research emphasises capability growth and tolerates low failure rates when average-case performance is high. AI safety and alignment resea

ERUnderstand: Evaluating Vision-Language Models on Structured ER Diagrams

Model ReleasesDGX agent

arXiv:2607.24707v1 Announce Type: new Abstract: Entity-Relationship Diagrams (ERDs) are central to conceptual database design, yet they are typically available only as rendered images rather than mach

Escaping the Euclidean Void: Manifold-Informed Flow Matching for Sequential Recommendation

Local AiDGX agent

arXiv:2607.23762v1 Announce Type: cross Abstract: Conventional recommenders capture users' preferences by optimizing observed user-item relations, whereas continuous generative recommendation addition

ESF-Bench: Benchmarking Challenging Slot-Filling Scenarios for Real-World Enterprise Applications

Model ReleasesDGX agent

arXiv:2607.23326v1 Announce Type: new Abstract: The rapid rise of large language models (LLMs) has driven transformative adoption across enterprises. However, deploying these models in real-world sett

ESRVS: Extreme Semi-Supervised Retinal Vessel Segmentation with a Single Annotated Image

ResearchDGX agent

arXiv:2607.24453v1 Announce Type: cross Abstract: Learning from minimal human supervision is a long-standing goal in medical image analysis, where dense expert annotations are costly. We study retinal

Evaluating and Mitigating the Misguidance Effect of Buggy Code in LLM-Generated Unit Tests

ResearchDGX agent

arXiv:2607.22883v1 Announce Type: cross Abstract: While Large Language Models (LLMs) show great promise for automating unit test generation, recent studies suggest that the quality of generated tests

Evaluating Large Language Models for Symbolic Security Protocol Analysis

Model ReleasesDGX agent

arXiv:2607.20712v1 Announce Type: cross Abstract: Security protocol verification relies on formal tools such as ProVerif and OFMC. This study evaluates whether Large Language Models (LLMs) can perform

Evaluating LLMs as Interpretable Controllers for Dynamical Systems

Model ReleasesDGX agent

arXiv:2607.22609v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for decision-making and reasoning tasks, yet their potential as controllers for physical systems rema

Evaluating RAG for French immigration law: a benchmark and baseline study

Model ReleasesDGX agent

arXiv:2607.24449v1 Announce Type: cross Abstract: International recruitment in France requires navigating a layered legal framework absent from existing legal AI benchmarks. We present a publicly avai

Evaluating the Impact of Explainable AI on Trust in AI-Assisted Code Review

ApplicationsDGX agent

arXiv:2607.24601v1 Announce Type: cross Abstract: Background: Large language models (LLMs) are increasingly used to automate code review, but the reasoning behind their decisions remains hard to under

Evaluating the Impact of Reviewer Guideline Design on LLM-Based Automated Peer Review

ResearchDGX agent

arXiv:2607.22553v1 Announce Type: cross Abstract: Peer review is an essential process in scientific research, yet the growing workload has made its automation increasingly necessary. In this study, we

EventOD: Event-Aware OD Flow Generation via LLM-Guided Semantic Modulation

ResearchDGX agent

arXiv:2607.22655v1 Announce Type: new Abstract: Estimating origin-destination (OD) flows under disruptive events is important for disaster response and urban resilience. Existing deep OD models traine

Every Client Is an Environment: Federated De-confounding for Spatio-Temporal Forecasting

ResearchDGX agent

arXiv:2607.24218v1 Announce Type: cross Abstract: Federated learning has emerged as a promising paradigm for spatio-temporal forecasting (STF), enabling collaborative model training without sharing ra

EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff

AgentsDGX agent

arXiv:2607.23955v1 Announce Type: new Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no c

Eviction as Estimation: A Fixed-Lag Smoothing View of Test-Time Memory, and When Measuring Beats Accumulating

SafetyDGX agent

arXiv:2607.24667v1 Announce Type: new Abstract: A language model with a bounded working memory must repeatedly decide which stored items to keep. Every deployed method decides the moment an item arriv

Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models

HardwareDGX agent

arXiv:2507.21438v2 Announce Type: replace Abstract: Ontologies and knowledge graphs require continuous evolution to remain comprehensive and accurate, but manual curation is labor intensive. Large Lan

Evolving from Lessons: Skill-Augmented Table Graph Reasoning for Operation-wise Table Question Answering

Model ReleasesDGX agent

arXiv:2607.22633v1 Announce Type: new Abstract: Table Question Answering (TableQA) aims to reason over tables to answer user queries. Existing research treats all questions uniformly and evaluates sol

Exact and Asymptotically Complete Robust Verifications of Neural Networks via Ising Solvers

ResearchDGX agent

arXiv:2603.00408v2 Announce Type: replace-cross Abstract: We present an Ising-compatible framework for formal neural-network robustness verification under bounded input perturbations. For piecewise-li

Exact values and exact upper bounds for families of integers with arithmetic progression intersections (Erdos Problem #272)

ResearchDGX agent

arXiv:2607.23004v1 Announce Type: cross Abstract: Let t(N) be the largest t for which there exist distinct sets A_1,ots,A_t subseteq {1,ots,N} such that A_i ap A_j is a nonempty arithmetic progression

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines

SafetyDGX agent

arXiv:2607.22569v1 Announce Type: new Abstract: Coding agents are increasingly integrated into system operations, where their tool use can directly modify project artifacts, execution environments, an

← Previous
1…4748495051…354
Next →