AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Tutorials

Leveraging generative models to assist Monte Carlo sampling

DGX agent

arXiv:2608.07648v1 Announce Type: cross Abstract: Sampling high-dimensional probability distributions is a central task in scientific computing, with applications ranging from Bayesian inference to st

tutorialsarxiv-cs-lg
11 Aug 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection

DGX agent

arXiv:2608.09593v1 Announce Type: cross Abstract: Recent advances in speech synthesis and audio generation have made high-fidelity acoustic forgery low-cost and difficult to attribute, enabling a real

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Towards Expressive and Faithful Audio-to-Image Generation: A Unified Multimodal Dataset and Synthesis Framework

DGX agent

arXiv:2608.09529v1 Announce Type: new Abstract: As an important subfield of cross-modal generation, synthesizing static visual content in the form of images from audio, namely audio-to-image (A2I) gen

safetyarxiv-cs-cv
11 Aug 2026
Agents

Autonomous discovery of accelerator commissioning algorithms

DGX agent

arXiv:2608.07138v1 Announce Type: cross Abstract: Simulated commissioning has become essential for de-risking modern light-source design and commissioning, but the procedures being simulated are still

agentsarxiv-cs-ai
10 Aug 2026
Safety

Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests

DGX agent

arXiv:2608.06908v1 Announce Type: cross Abstract: We propose Zero-phase Component Analysis (ZCA) whitening as a geometric pre-processing step for the Word Embedding Association Test (WEAT). WEAT is a

safetyarxiv-cs-ai
10 Aug 2026
Safety

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

DGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

I Seek You in Videos: Identity-Conditioned Queries for Person-Centric Video Reasoning

DGX agent

arXiv:2608.07417v1 Announce Type: cross Abstract: Real-world video reasoning often involves multimodal, multi-source inputs, whereas existing video reasoning tasks typically assume a simplified video-

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

MAC: A Conversion Rate Prediction Benchmark Featuring Labels Under Multiple Attribution Mechanisms

DGX agent

arXiv:2603.02184v2 Announce Type: replace-cross Abstract: Multi-attribution learning (MAL), which enhances model performance by learning from conversion labels yielded by multiple attribution mechanis

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery

DGX agent

arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether they can support complex real laboratory

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Vehicle routing problem using deep reinforcement learning - A case study about truck planning in the industry

DGX agent

arXiv:2608.06668v1 Announce Type: new Abstract: As an important component of the supply chain industry, transportation has experienced rapid development in the past decade with the assistance of digit

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader

DGX agent

arXiv:2608.06474v1 Announce Type: new Abstract: Large language models increasingly generate complete websites from natural-language descriptions, and reinforcement learning has become a central approa

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

DGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

DGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset

DGX agent

arXiv:2608.06165v1 Announce Type: cross Abstract: Existing audio-to-score (A2S) systems primarily focus on classical music, and the application to popular music remains underexplored. This paper first

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for Political News Evaluation

DGX agent

arXiv:2608.05155v1 Announce Type: cross Abstract: Traditional sentiment analysis (SA) models, while effective for polarity classification, provide limited insight into the rhetorical, ideological, and

safetyarxiv-cs-ai
7 Aug 2026
Safety

FI-TW: An Open Train-Weather Dataset for Railway Delay Analysis in Finland

DGX agent

arXiv:2601.16592v2 Announce Type: replace-cross Abstract: Train delays result from complex interactions between operational, technical, and environmental factors. While weather impacts railway reliabi

safetyarxiv-cs-ai
7 Aug 2026
Applications

Investigating Artificial Intelligence Digital Sovereignty in Mobile Shopping Apps: A Case Study of Nigeria

DGX agent

arXiv:2608.06364v1 Announce Type: cross Abstract: The use of e-commerce mobile applications is expanding in Nigeria, creating both opportunities and risks, including fraud and reduced user control ove

applicationsarxiv-cs-ai
7 Aug 2026
Safety

Large Language Models Threaten Double-blind Review

DGX agent

arXiv:2608.05157v1 Announce Type: cross Abstract: Double blind peer review serves as the scientific community primary defense against status and affiliation bias. Its effectiveness rests on the assump

safetyarxiv-cs-ai
7 Aug 2026
Agents

CheMLFlow: An Open-Source Platform for Cheminformatics and Materials Informatics Applications

DGX agent

arXiv:2608.04942v1 Announce Type: cross Abstract: CheMLFlow is an open-source platform for building and executing end-to-end, high-throughput, and agentic workflows for scientific and technological ap

agentsarxiv-cs-ai
6 Aug 2026
Agents

Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings for CLEF JOKER 2025 Task 2

DGX agent

arXiv:2507.06506v2 Announce Type: replace-cross Abstract: Translating wordplay across languages presents unique challenges that have long confounded both professional human translators and machine tra

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

SimMOF: AI agent for Automated MOF Simulations

DGX agent

arXiv:2603.29152v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) offer a vast design space, and as such, computational simulations play a critical role in predicting their structura

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

A Unified Framework for Human AI Collaboration in Security Operations Centers with Trusted Autonomy

DGX agent

arXiv:2505.23397v3 Announce Type: replace Abstract: This article presents a structured framework for Human-AI collaboration in Security Operations Centers (SOCs), integrating AI autonomy, trust calibr

agentsarxiv-cs-ai
5 Aug 2026
Applications

Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems

DGX agent

arXiv:2608.03413v1 Announce Type: new Abstract: As artificial intelligence (AI) continues to evolve and mature, recent AI practices have moved beyond large language models (LLMs) and text or image gen

applicationsarxiv-cs-ai
5 Aug 2026
Hardware

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

DGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

hardwarearxiv-cs-ai
5 Aug 2026
Model Releases

Looking under the Wrong Lamppost: On the Limitations of Automated Translation Quality Estimation

DGX agent

arXiv:2608.03577v1 Announce Type: new Abstract: Automation of Translation Quality Estimation (QE) has emerged as a widely discussed approach to managing translation quality at scale, and a growing num

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

MDArena: Evaluating Coding Agents on Realistic Molecular Dynamics Workflows

DGX agent

arXiv:2608.02642v1 Announce Type: cross Abstract: Accelerating scientific discovery is among the most consequential applications of AI, and computational biomolecular simulation stands out as a partic

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ArabicDialectSafety: A Dialect-Aware Benchmark for Arabic Content Safety Classification

DGX agent

arXiv:2608.01291v1 Announce Type: new Abstract: We present ArabicDialectSafety, a human-curated Arabic safety dataset of 25,071 prompts covering six Arabic varieties: Modern Standard Arabic, Syrian, E

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

DGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Style Wins, Substance Loses: A Diagnosis of LLM-as-Judge in Idea Generation

DGX agent

arXiv:2608.01666v1 Announce Type: new Abstract: However, whether these judges truly evaluate the scientific substance of ideas or are influenced by superficial stylistic presentation remains an open q

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

AgentHPOBench: A Benchmark For Evaluating LLM Agents as Sequential Hyperparameter Optimizers

DGX agent

arXiv:2607.29626v1 Announce Type: new Abstract: As LLMs evolve from code completion systems into autonomous scientific agents, evaluating their ability to conduct experiments has become increasingly i

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Benchmarks Are Not Validation: A System-Level View of Financial LLM Applications

DGX agent

arXiv:2607.28840v1 Announce Type: new Abstract: Large language models are increasingly deployed in financial applications that combine retrieval, proprietary data, tool use, orchestration logic, monit

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

CBCT-IQ: A Publicly Available Annotated Cone-Beam CT Dataset for Image Quality Assessment and Benchmarking

DGX agent

arXiv:2607.29253v1 Announce Type: cross Abstract: Medical image quality plays a critical role in diagnostic accuracy, especially in X-ray-based imaging modalities such as cone-beam computed tomography

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

DGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

A Distributed Acoustic Sensing Dataset for Vessel Detection and Localization in Submarine Cable Protection

DGX agent

arXiv:2607.28306v1 Announce Type: cross Abstract: Recent incidents of accidental damage and suspected sabotage to submarine telecommunication and power cables, particularly in the Baltic Sea, have und

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

DGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

model-releasesarxiv-cs-cl
31 Jul 2026
Applications

HealthCAT: An Interpretable Encoder-only Transformer Framework for Health Indicator Prediction and Temporal Interpretation of Wearable Sensor Data

DGX agent

arXiv:2607.27635v1 Announce Type: cross Abstract: Wearable sensors continuously capture fine-grained multivariate time-series data, providing opportunities to model behavioural patterns associated wit

applicationsarxiv-cs-lg
31 Jul 2026
Model Releases

IDP AutoOpt: Agent-Driven Optimization of Document Processing Pipeline Configurations

DGX agent

arXiv:2607.26075v1 Announce Type: cross Abstract: We present IDP AutoOpt, an autonomous LLM agent that discovers high-performing configurations for intelligent document processing (IDP) pipelines. Tun

model-releasesarxiv-cs-ai
31 Jul 2026
Agents

Multi-Agent Debate Strategies: Survey, Taxonomy, and Challenges

DGX agent

arXiv:2607.26212v1 Announce Type: cross Abstract: Multi-Agent Debate (MAD) is a promising paradigm for improving the accuracy and robustness of Large Language Model (LLM)-based agentic systems. It ena

agentsarxiv-cs-ai
31 Jul 2026
Agents

Partner Capability Estimation for Task-Agnostic Adaptation in Ad-Hoc Teamwork

DGX agent

arXiv:2607.27177v1 Announce Type: new Abstract: Effective collaboration with novel and diverse partners is a crucial skill for autonomous agents. Most current ad-hoc teamwork (AHT) approaches assume t

agentsarxiv-cs-ai
31 Jul 2026
Tutorials

The Social Cost of an AI Teammate: How an Artificial Teammate Reshapes Human-Human Communication in Small-Team Decision-Making

DGX agent

arXiv:2607.27179v1 Announce Type: cross Abstract: Conversational AI is increasingly positioned as a teammate rather than a tool, yet we know little about how its presence reshapes communication among

tutorialsarxiv-cs-ai
31 Jul 2026
Local Ai

When Should AI Follow? Task Structure and Joint Adaptation by Human and AI Agents

DGX agent

arXiv:2504.20903v4 Announce Type: replace-cross Abstract: How should organizations divide and sequence decision tasks between human and artificial agents? We develop a computational model of joint seq

local-aiarxiv-cs-ai
31 Jul 2026
Applications

Neural Radiance Fields for the Real World: A Survey

DGX agent

arXiv:2501.13104v3 Announce Type: replace Abstract: Neural Radiance Fields (NeRFs) have remodeled 3D scene representation since release. NeRFs can effectively reconstruct complex 3D scenes from 2D ima

applicationsarxiv-cs-cv
30 Jul 2026
Model Releases

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

DGX agent

arXiv:2607.26873v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) enables language models to self-evolve at inference time without labeled feedback. Existing methods rely on answ

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications

DGX agent

arXiv:2607.25642v1 Announce Type: cross Abstract: Instruction-based Image Editing (IIE) aims to transform a given image into a new one based on textual instructions. Advances in Large Language Models

model-releasesarxiv-cs-cl
29 Jul 2026
Agents

Agentic Autoresearch for CT Reconstruction

DGX agent

arXiv:2607.22824v1 Announce Type: cross Abstract: Comparing CT reconstruction methods fairly is labor-intensive and largely manual, and many benchmarks use idealized data. We ask whether a large langu

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Algorithmic Blindness in Large Language Models: A Calibration Study of Performance Prediction

DGX agent

arXiv:2602.21947v5 Announce Type: replace Abstract: Large language models (LLMs) demonstrate remarkable breadth of knowledge, yet their ability to reason about computational processes remains poorly u

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

AutoMat: Enabling Automated Crystal Structure Reconstruction from Microscopy via Agentic Tool Use

DGX agent

arXiv:2505.12650v2 Announce Type: replace-cross Abstract: Reconstructing atomistic crystal structures from a single noisy STEM projection is an ill-posed inverse problem: multiple lattices can explain

model-releasesarxiv-cs-ai
28 Jul 2026
Hardware

Energy Constrained Hierarchical Underwater Monitoring via Local Multi-Agent RAG

DGX agent

arXiv:2607.24313v1 Announce Type: cross Abstract: Marine life monitoring is limited by strict energy constraints, poor underwater connectivity, and the high cost of transmitting raw multimodal data fr

hardwarearxiv-cs-cv
28 Jul 2026
← Previous
1…401402403404405…466
Next →