AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

QuantaAlpha: An Evolutionary Framework for LLM-Driven Alpha Mining

DGX agent

arXiv:2602.07085v2 Announce Type: replace-cross Abstract: Financial markets are noisy and non-stationary, making alpha mining highly sensitive to noise in backtesting results and sudden market regime

model-releasesarxiv-cs-ai
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Querying Inconsistent Prioritized Data with ORBITS: Algorithms, Implementation, and Experiments

DGX agent

arXiv:2202.07980v4 Announce Type: replace-cross Abstract: We investigate practical algorithms for inconsistency-tolerant query answering over prioritized knowledge bases, which consist of a logical th

researcharxiv-cs-ai
23 Apr 2026
Research

Rabies diagnosis in low-data settings: A comparative study on the impact of data augmentation and transfer learning

DGX agent

arXiv:2604.19823v1 Announce Type: cross Abstract: Rabies remains a major public health concern across many African and Asian countries, where accurate diagnosis is critical for effective epidemiologic

researcharxiv-cs-ai
23 Apr 2026
Model Releases

ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability

DGX agent

arXiv:2508.07050v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of Large

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Recency Biased Causal Attention for Time-series Forecasting

DGX agent

arXiv:2502.06151v2 Announce Type: replace-cross Abstract: Recency bias is a useful inductive prior for sequential modeling: it emphasizes nearby observations and can still allow longer-range dependenc

safetyarxiv-cs-ai
23 Apr 2026
Safety

Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem

DGX agent

arXiv:2604.20805v1 Announce Type: cross Abstract: The value alignment problem for artificial intelligence (AI) is often framed as a purely technical or normative challenge, sometimes focused on hypoth

safetyarxiv-cs-ai
23 Apr 2026
Safety

Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport

DGX agent

arXiv:2510.01706v2 Announce Type: replace-cross Abstract: Standard representational similarity methods align each layer of a network to its best match in another independently, producing asymmetric re

safetyarxiv-cs-ai
23 Apr 2026
Safety

Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model

DGX agent

arXiv:2604.19838v1 Announce Type: new Abstract: Understanding how road users resolve space-sharing conflicts is important both for traffic safety and the safe deployment of autonomous vehicles. While

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

RSRCC: A Remote Sensing Regional Change Comprehension Benchmark Constructed via Retrieval-Augmented Best-of-N Ranking

DGX agent

arXiv:2604.20623v1 Announce Type: cross Abstract: Traditional change detection identifies where changes occur, but does not explain what changed in natural language. Existing remote sensing change cap

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Same Content, Different Answers: Cross-Modal Inconsistency in MLLMs

DGX agent

arXiv:2512.08923v2 Announce Type: replace Abstract: We introduce two new benchmarks REST and REST+ (Render-Equivalence Stress Tests) to enable systematic evaluation of cross-modal inconsistency in mul

researcharxiv-cs-ai
23 Apr 2026
Research

Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models

DGX agent

arXiv:2604.19768v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic miscalibration with rhetorical intensity not proportionate to epistemic grounding. This study tests th

researcharxiv-cs-ai
23 Apr 2026
Applications

Scalable AI Inference: Performance Analysis and Optimization of AI Model Serving

DGX agent

arXiv:2604.20420v1 Announce Type: cross Abstract: AI research often emphasizes model design and algorithmic performance, while deployment and inference remain comparatively underexplored despite being

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

SciCoQA: Quality Assurance for Scientific Paper--Code Alignment

DGX agent

arXiv:2601.12910v3 Announce Type: replace-cross Abstract: Discrepancies between scientific papers and their code undermine reproducibility, a concern that grows as automated research agents scale scie

model-releasesarxiv-cs-ai
23 Apr 2026
Research

scpFormer: A Foundation Model for Unified Representation and Integration of the Single-Cell Proteomics

DGX agent

arXiv:2604.20003v1 Announce Type: cross Abstract: The integration of single-cell proteomic data is often hindered by the fragmented nature of targeted antibody panels. To address this limitation, we i

researcharxiv-cs-ai
23 Apr 2026
Applications

Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction

DGX agent

arXiv:2604.20311v1 Announce Type: cross Abstract: Micro-video popularity prediction (MVPP) aims to forecast the future popularity of videos on online media, which is essential for applications such as

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness

DGX agent

arXiv:2604.20413v1 Announce Type: new Abstract: Large language models perform well on many reasoning tasks, yet they often lack awareness of whether their current knowledge or reasoning state is compl

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Self-Describing Structured Data with Dual-Layer Guidance: A Lightweight Alternative to RAG for Precision Retrieval in Large-Scale LLM Knowledge Navigation

DGX agent

arXiv:2604.19777v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit a well-documented positional bias when processing long input contexts: information in the middle of a context win

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning

DGX agent

arXiv:2604.20601v1 Announce Type: new Abstract: We introduce SuperIgor, a framework for instruction-following tasks. Unlike prior methods that rely on predefined subtasks, SuperIgor enables a language

agentsarxiv-cs-ai
23 Apr 2026
Safety

Semantic Prompting: Agentic Incremental Narrative Refinement through Spatial Semantic Interaction

DGX agent

arXiv:2604.19971v1 Announce Type: cross Abstract: Interactive spatial layouts empower users to synthesize information and organize findings for sensemaking. While Large Language Models (LLMs) can auto

safetyarxiv-cs-ai
23 Apr 2026
Research

Semantic Recall for Vector Search

DGX agent

arXiv:2604.20417v1 Announce Type: cross Abstract: We introduce Semantic Recall, a novel metric to assess the quality of approximate nearest neighbor search algorithms by considering only semantically

researcharxiv-cs-ai
23 Apr 2026
Research

Separable Pathways for Causal Reasoning: How Architectural Scaffolding Enables Hypothesis-Space Restructuring in LLM Agents

DGX agent

arXiv:2604.20039v1 Announce Type: new Abstract: Causal discovery through experimentation and intervention is fundamental to robust problem solving. It requires not just updating beliefs within a fixed

researcharxiv-cs-ai
23 Apr 2026
Agents

Shift-Up: A Framework for Software Engineering Guardrails in AI-native Software Development -- Initial Findings

DGX agent

arXiv:2604.20436v1 Announce Type: cross Abstract: Generative AI (GenAI) is reshaping software engineering by shifting development from manual coding toward agent-driven implementation. While vibe codi

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation

DGX agent

arXiv:2604.19793v1 Announce Type: new Abstract: LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Skyline-First Traversal as a Control Mechanism for Multi-Criteria Graph Search

DGX agent

arXiv:2604.19807v1 Announce Type: new Abstract: In multi-criteria graph traversal, paths are compared via Pareto dominance, an ordering that identifies which paths are non-dominated, but says nothing

researcharxiv-cs-ai
23 Apr 2026
Model Releases

SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models

DGX agent

arXiv:2509.15174v3 Announce Type: replace-cross Abstract: WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMAR

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Soft-Label Governance for Distributional Safety in Multi-Agent Systems

DGX agent

arXiv:2604.19752v1 Announce Type: cross Abstract: Multi-agent AI systems exhibit emergent risks that no single agent produces in isolation. Existing safety frameworks rely on binary classifications of

model-releasesarxiv-cs-ai
23 Apr 2026
Research

SolidCoder: Bridging the Mental-Reality Gap in LLM Code Generation through Concrete Execution

DGX agent

arXiv:2604.19825v1 Announce Type: cross Abstract: State-of-the-art code generation frameworks rely on mental simulation, where LLMs internally trace execution to verify correctness. We expose a fundam

researcharxiv-cs-ai
23 Apr 2026
Model Releases

SpeechParaling-Bench: A Comprehensive Benchmark for Paralinguistic-Aware Speech Generation

DGX agent

arXiv:2604.20842v1 Announce Type: cross Abstract: Paralinguistic cues are essential for natural human-computer interaction, yet their evaluation in Large Audio-Language Models (LALMs) remains limited

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

SphUnc: Hyperspherical Uncertainty Decomposition and Causal Identification via Information Geometry

DGX agent

arXiv:2603.01168v2 Announce Type: replace-cross Abstract: Reliable decision-making in complex multi-agent systems requires calibrated predictions and interpretable uncertainty. We introduce SphUnc, a

agentsarxiv-cs-ai
23 Apr 2026
Research

Stabilising Generative Models of Attitude Change

DGX agent

arXiv:2604.19791v1 Announce Type: new Abstract: Attitude change - the process by which individuals revise their evaluative stances - has been explained by a set of influential but competing verbal the

researcharxiv-cs-ai
23 Apr 2026
Applications

Stateless Decision Memory for Enterprise AI Agents

DGX agent

arXiv:2604.20158v1 Announce Type: new Abstract: Enterprise deployment of long-horizon decision agents in regulated domains (underwriting, claims adjudication, tax examination) is dominated by retrieva

applicationsarxiv-cs-ai
23 Apr 2026
Agents

Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine

DGX agent

arXiv:2604.20022v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous diagnostic agents, yet they conflate two fundamentally different capabilities: natural-l

agentsarxiv-cs-ai
23 Apr 2026
Safety

Storm Surge Modeling, Bias Correction, Graph Neural Networks, Graph Convolution Networks

DGX agent

arXiv:2604.20688v1 Announce Type: cross Abstract: Storm surge forecasting remains a critical challenge in mitigating the impacts of tropical cyclones on coastal regions, particularly given recent tren

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Supplement Generation Training for Enhancing Agentic Task Performance

DGX agent

arXiv:2604.20727v1 Announce Type: cross Abstract: Training large foundation models for agentic tasks is increasingly impractical due to the high computational costs, long iteration cycles, and rapid o

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Surrogate modeling for interpreting black-box LLMs in medical predictions

DGX agent

arXiv:2604.20331v1 Announce Type: cross Abstract: Large language models (LLMs), trained on vast datasets, encode extensive real-world knowledge within their parameters, yet their black-box nature obsc

applicationsarxiv-cs-ai
23 Apr 2026
Agents

SWE-chat: Coding Agent Interactions From Real Users in the Wild

DGX agent

arXiv:2604.20779v1 Announce Type: new Abstract: AI coding agents are being adopted at scale, yet we lack empirical evidence on how people actually use them and how much of their output is useful in pr

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

SweRank: Software Issue Localization with Code Ranking

DGX agent

arXiv:2505.07849v2 Announce Type: replace-cross Abstract: Software issue localization, the task of identifying the precise code locations (files, classes, or functions) relevant to a natural language

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Taint-Style Vulnerability Detection and Confirmation for Node.js Packages Using LLM Agent Reasoning

DGX agent

arXiv:2604.20179v1 Announce Type: cross Abstract: The rapidly evolving Node.js ecosystem currently includes millions of packages and is a critical part of modern software supply chains, making vulnera

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models

DGX agent

arXiv:2508.18609v4 Announce Type: replace-cross Abstract: Post-Training Quantization (PTQ) is a critical strategy for efficient Large Language Models (LLMs) deployment. However, existing scaling laws

researcharxiv-cs-ai
23 Apr 2026
Research

Text Steganography with Dynamic Codebook and Multimodal Large Language Model

DGX agent

arXiv:2604.20269v1 Announce Type: cross Abstract: With the popularity of the large language models (LLMs), text steganography has achieved remarkable performance. However, existing methods still have

researcharxiv-cs-ai
23 Apr 2026
Research

Text to model via SysML: Automated generation of dynamical system computational models from unstructured natural language text via enhanced System Modeling Language diagrams

DGX agent

arXiv:2507.06803v3 Announce Type: replace-cross Abstract: This paper contributes to speeding up the design and deployment of engineering dynamical systems by proposing a strategy for exploiting domain

researcharxiv-cs-ai
23 Apr 2026
Agents

The AI Telco Engineer: Toward Autonomous Discovery of Wireless Communications Algorithms

DGX agent

arXiv:2604.19803v1 Announce Type: new Abstract: Agentic AI is rapidly transforming the way research is conducted, from prototyping ideas to reproducing results found in the literature. In this paper,

agentsarxiv-cs-ai
23 Apr 2026
Safety

The Existential Theory of Research: Why Discovery Is Hard

DGX agent

arXiv:2604.19810v1 Announce Type: new Abstract: Can scientific discovery be made arbitrarily easy by choosing the right representation, collecting enough data, and deploying sufficiently powerful algo

safetyarxiv-cs-ai
23 Apr 2026
Research

The Expense of Seeing: Attaining Trustworthy Multimodal Reasoning Within the Monolithic Paradigm

DGX agent

arXiv:2604.20665v1 Announce Type: cross Abstract: The rapid proliferation of Vision-Language Models (VLMs) is widely celebrated as the dawn of unified multimodal knowledge discovery but its foundation

researcharxiv-cs-ai
23 Apr 2026
Model Releases

The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning

DGX agent

arXiv:2603.29025v2 Announce Type: replace-cross Abstract: Large language models systematically fail when a salient surface cue conflicts with an unstated feasibility constraint. We study this through

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

The OpenHands Software Agent SDK: A Composable and Extensible Foundation for Production Agents

DGX agent

arXiv:2511.03690v2 Announce Type: replace-cross Abstract: Agents are now used widely in the process of software development, but building production-ready software engineering agents is a complex task

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

DGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?

DGX agent

arXiv:2604.19749v1 Announce Type: new Abstract: Equipping LLMs with external tools effectively addresses internal reasoning limitations. However, it introduces a critical yet under-explored phenomenon

safetyarxiv-cs-ai
23 Apr 2026
← Previous
1…393394395396397…443
Next →