AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Model Releases

SalAngaBhava: A Sinhala Market Dataset for Aspect-based Sentiment Analysis

DGX agent

arXiv:2607.05259v1 Announce Type: new Abstract: Sentiment analysis has been a primary domain under Natural Language Processing (NLP) from its inception as it plays a vital role in both real-world and

model-releasesarxiv-cs-cl
7 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Teaming Up with AI: Coordination and Cooperation

DGX agent

arXiv:2607.03181v1 Announce Type: cross Abstract: Successful diffusion of AI in the workforce hinges on the economic value that AI brings to human endeavors. Bringing AI into the workforce is more tha

safetyarxiv-cs-ai
7 Jul 2026
Safety

The Foreign Policy AI Evaluation Gap

DGX agent

arXiv:2607.02955v1 Announce Type: cross Abstract: We argue that AI systems used in conducting foreign policy tasks - broadly enacting 'statecraft' - should be a priority test case for technical AI gov

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

DGX agent

arXiv:2407.11691v5 Announce Type: replace Abstract: We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friend

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

From Approximation to Emergence: A Theory of Deep Learning

DGX agent

arXiv:2607.01311v1 Announce Type: new Abstract: Deep learning has outgrown any single mathematical explanation. From Approximation to Emergence develops a unified, proof-oriented account of modern dee

safetyarxiv-cs-lg
3 Jul 2026
Model Releases

PreScience: A Dataset and Benchmark for Scientific Forecasting

DGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Conversable Complexity: Agentic LLM Collectives as Interpretable Substrates

DGX agent

arXiv:2607.01047v1 Announce Type: new Abstract: Complexity and interpretability rarely coincide: systems rich enough for complex behaviours to emerge are usually too opaque to question, while transpar

agentsarxiv-cs-cl
2 Jul 2026
Model Releases

Quantum vs. Classical Machine Learning: A Unified Empirical Comparison

DGX agent

arXiv:2607.01197v1 Announce Type: new Abstract: Quantum computing has emerged as a promising computational paradigm for machine learning (ML), with the potential to offer computational advantages over

model-releasesarxiv-cs-lg
2 Jul 2026
Agents

Rethinking Multi-Label Image Classification With Deep Learning: Taxonomy, Challenge, and Outlook

DGX agent

arXiv:2607.00839v1 Announce Type: new Abstract: Multi-label image classification (MLIC), a fundamental task in computer vision, focuses on identifying multiple objects or concepts within an image, und

agentsarxiv-cs-cv
2 Jul 2026
Model Releases

Sheet Music Benchmark: Standardized Optical Music Recognition Evaluation

DGX agent

arXiv:2506.10488v3 Announce Type: replace Abstract: In this work, we introduce the Sheet Music Benchmark (SMB), a dataset of six hundred and eighty-five pages specifically designed to benchmark Optica

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Svarna: An Open Corpus Workbench for Modern Greek

DGX agent

arXiv:2607.00970v1 Announce Type: new Abstract: This paper introduces Svarna, a free, open-source, web-based corpus workbench for modern Greek. Svarna integrates five databases covering various regist

model-releasesarxiv-cs-cl
2 Jul 2026
Safety

FLARE-AI: Flaw Reporting for AI

DGX agent

arXiv:2606.31567v1 Announce Type: cross Abstract: Flaw reporting for deployed AI systems is fundamental to identifying system failures and improving AI safety. Yet the AI reporting ecosystem is fragme

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

Accelerating scientific discovery with Co-Scientist

DGX agent

arXiv:2502.18864v2 Announce Type: replace Abstract: Scientific discovery is driven by scientists generating novel hypotheses for complex problems that undergo rigorous experimental validation. To augm

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Are Humans Evolved Instruction Followers? An Underlying Inductive Bias Enables Rapid Instructed Task Learning

DGX agent

arXiv:2606.29792v1 Announce Type: new Abstract: Human adults can often perform a novel task correctly on the first attempt after only receiving verbal or written instructions. This rapid instructed ta

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios?

DGX agent

arXiv:2606.29920v1 Announce Type: new Abstract: Rubric-based scoring has become a widely used paradigm in model evaluation, typically with LLM-as-a-Judge (LaaJ) for rubric scoring. However, the reliab

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

RoAd-RL: A Unified Library and Benchmark for Robust Adversarial Reinforcement Learning

DGX agent

arXiv:2606.29867v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in robotics and autonomous systems, yet remains vulnerable to adversarial perturbat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TextClusterLab: An Integrated Framework for Reliable Text Clustering Studies

DGX agent

arXiv:2606.28328v1 Announce Type: cross Abstract: In recent years, text clustering has become a critical technique for applications including intent discovery, topic mining, and recommendation systems

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

An Empirical Analysis of Factual Errors in Human-Written Text and its Application

DGX agent

arXiv:2606.27959v1 Announce Type: new Abstract: Factual Error Detection (FED), which is the task of identifying factually incorrect spans in a given text, has long been recognized as an important rese

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

QuantV2X: A Fully Quantized Multi-Agent System for Cooperative Perception

DGX agent

arXiv:2509.03704v2 Announce Type: replace Abstract: Cooperative perception through Vehicle-to-Everything (V2X) communication offers significant potential for enhancing vehicle perception by mitigating

model-releasesarxiv-cs-cv
29 Jun 2026
Tutorials

Neural Architecture Search for Generative Adversarial Networks: A Comprehensive Review and Critical Analysis

DGX agent

arXiv:2606.26169v1 Announce Type: cross Abstract: Neural Architecture Search (NAS) has emerged as a pivotal technique in optimizing the design of Generative Adversarial Networks (GANs), automating the

tutorialsarxiv-cs-ai
26 Jun 2026
Applications

LLM Evolution as an Industry-Scale Ecosystem: A Lifecycle Perspective on Continual Learning

DGX agent

arXiv:2606.24901v1 Announce Type: new Abstract: Continual learning capability is critical for Industrial LLMs, as deployed models must be continuously updated to meet evolving requirements and environ

applicationsarxiv-cs-lg
25 Jun 2026
Agents

SwarmFly: A simulation platform for UAV swarm experiment design and validation

DGX agent

arXiv:2606.25146v1 Announce Type: new Abstract: The initial development phase of UAV swarms largely depends on simulation for experimental design and validation, yet existing open-source tools are oft

agentsarxiv-cs-ro
25 Jun 2026
Model Releases

A Benchmark of State-Space Models vs. Transformers and BiLSTM-based Models for Historical Newspaper OCR

DGX agent

arXiv:2604.00725v2 Announce Type: replace Abstract: End-to-end OCR for historical newspapers remains challenging, as models must handle long text sequences, degraded print quality, and complex layouts

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

A Geometry-Informed Computer Vision Method for Detecting and Examining Overtaking Vehicles From A Bicycle

DGX agent

arXiv:2606.23699v1 Announce Type: new Abstract: Instrumented bicycle studies have produced direct field evidence on vehicle passing behavior, but extracting overtaking events from continuous rear-faci

safetyarxiv-cs-cv
24 Jun 2026
Model Releases

AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approach

DGX agent

arXiv:2606.24655v1 Announce Type: cross Abstract: The explosive growth and complexity of product data within the dynamic Brazilian e-commerce landscape demand robust and specialized methods for struct

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

KANLib -- A Modular, Extensible and Fast Kolmogorov-Arnold Network Implementation

DGX agent

arXiv:2606.17927v2 Announce Type: replace-cross Abstract: Kolmogorov-Arnold Networks (KANs) have recently emerged as a promising alternative to traditional multilayer perceptrons by replacing linear w

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Quant Convergence: Bridging Classical Value Investing and Modern Factor Models for Systematic Equity Selection

DGX agent

arXiv:2606.24575v1 Announce Type: new Abstract: Modern finance relies heavily on complex machine learning models to find patterns in the stock market. However, as these AI models get more complicated,

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

A-Evolve-Training: Autonomous Post-Training of a 30B Model

DGX agent

arXiv:2606.20657v1 Announce Type: cross Abstract: Post-training a frontier model is normally weeks of human work: proposing data and recipe changes, launching runs, reading evals, deciding what to kee

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

ARVO: Atlas of Reproducible Vulnerabilities for Open-Source Software

DGX agent

arXiv:2408.02153v2 Announce Type: replace-cross Abstract: Achieving reproducibility, quantity, and diversity in vulnerability datasets has long been viewed as an inherent three-way trade-off, where im

applicationsarxiv-cs-lg
23 Jun 2026
Safety

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

DGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

safetyarxiv-cs-cv
23 Jun 2026
Safety

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning

DGX agent

arXiv:2606.21943v1 Announce Type: new Abstract: Reinforcement learning (RL) has become central to LLM post-training, yet the methods that dominate current pipelines, PPO and GRPO, represent only a nar

safetyarxiv-cs-lg
23 Jun 2026
Safety

Overcoming Imperfect Kinematics in Surgical Robotics Through Sim-to-Real Visuomotor Learning

DGX agent

arXiv:2606.21396v1 Announce Type: new Abstract: Robot-Assisted Surgery is integral to modern minimally invasive procedures, with automation emerging as the next frontier to enhance precision and reduc

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

RouteJudge: An Open Platform for Reproducible and Preference-Aware LLM Routing

DGX agent

arXiv:2606.18774v2 Announce Type: replace Abstract: We present RouteJudge, an online pairwise preference evaluation framework for LLM routing systems, with a public platform available at https://route

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

DGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Internet of Everything in the 6G Era: Paradigms, Enablers, Potentials and Future Directions

DGX agent

arXiv:2604.25018v2 Announce Type: replace-cross Abstract: The Internet of Everything (IoE) represents an evolution of the Internet of Things (IoT) by integrating people, data, processes, and things in

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

On the Limits of LLM-as-Judge for Scientific Novelty Assessment

DGX agent

arXiv:2606.12071v1 Announce Type: cross Abstract: LLMs are increasingly used to generate and judge scientific ideas. This makes novelty evaluation a central problem. Full idea evaluation is difficult

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes

DGX agent

arXiv:2606.11470v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across natural language processing tasks, yet reliable reasoning remains an open challenge

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

DGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

FreshRetailNet-LT: A Stockout-Annotated Censored Demand Dataset for Latent Demand Recovery and Forecasting in Fresh Retail

DGX agent

arXiv:2505.16319v3 Announce Type: replace Abstract: Accurate demand estimation is critical for the retail business in guiding the inventory and pricing policies of perishable products. However, it fac

model-releasesarxiv-cs-lg
10 Jun 2026
Tutorials

Integrated Real-Time Motion Tracking and AI Analysis for Athletic Performance Optimization

DGX agent

arXiv:2606.09842v1 Announce Type: cross Abstract: Applying Human Pose Estimation (HPE) in real world environments remains a challenging task, this paper explores and surveys real time HPE approaches a

tutorialsarxiv-cs-ai
10 Jun 2026
Agents

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving

DGX agent

arXiv:2602.23499v4 Announce Type: replace-cross Abstract: Collecting a high-quality dataset is a critical task that demands meticulous attention to detail, as overlooking certain aspects can render th

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

The 1st PortraitCraft Challenge: A CVPR 2026 Workshop Competition on Portrait Composition Understanding and Generation

DGX agent

arXiv:2606.10894v1 Announce Type: new Abstract: This paper presents an overview of the inaugural PortraitCraft Challenge, held as one of the official competitions at CVPR 2026. The challenge focuses o

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

A Survey of Heterogeneous Graph Neural Networks for Cybersecurity Anomaly Detection

DGX agent

arXiv:2510.26307v3 Announce Type: replace-cross Abstract: Anomaly detection is a critical task in cybersecurity, where identifying insider threats, access violations, and coordinated attacks is essent

model-releasesarxiv-cs-lg
9 Jun 2026
Agents

A Survey on Deep Multi-Task Learning in Connected Autonomous Vehicles

DGX agent

arXiv:2508.00917v2 Announce Type: replace-cross Abstract: Connected autonomous vehicles (CAVs) must simultaneously perform multiple tasks, such as perception, prediction, planning, and control, to ens

agentsarxiv-cs-cv
9 Jun 2026
Safety

AeroSpectra Sentinel: An Auditable LLM Prompt-Chaining Decision-Support Workflow for Acute Asthma Risk Assessment from Respiratory Sounds and Clinical Signals

DGX agent

arXiv:2606.08247v1 Announce Type: cross Abstract: Acute asthma risk assessment requires rapid interpretation of respiratory sounds, oxygenation, airflow limitation, speech ability, work of breathing,

safetyarxiv-cs-ai
9 Jun 2026
Agents

Beyond Agent Architecture: Execution Assumptions and Reproducibility in LLM-Based Trading Systems

DGX agent

arXiv:2606.08285v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems are increasingly proposed for financial trading, yet their reported performance remains difficult to co

agentsarxiv-cs-ai
9 Jun 2026
Safety

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation

DGX agent

arXiv:2606.09639v1 Announce Type: new Abstract: The fidelity and structural diversity of training datasets fundamentally determine the capabilities of video generation models. While commercial systems

safetyarxiv-cs-cv
9 Jun 2026
← Previous
1…403404405406407…466
Next →