AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Model Releases

McNdroid: A Longitudinal Multimodal Benchmark for Robust Drift Detection in Android Malware

DGX agent

arXiv:2605.06894v1 Announce Type: cross Abstract: Machine learning (ML) in real-world systems must contend with concept drift, adversarial actors, and a spectrum of potential features with varying cos

model-releasesarxiv-cs-lg
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants

DGX agent

arXiv:2603.09652v3 Announce Type: replace Abstract: With the rapid advancement of Large Language Models (LLMs) in code generation, human-AI interaction is evolving from static text responses to dynami

model-releasesarxiv-cs-ai
11 May 2026
Applications

MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes

DGX agent

arXiv:2605.06897v1 Announce Type: cross Abstract: The rise of Internet of Things (IoT) devices in the physical world necessitates voice-based interfaces capable of handling complex user experiences. W

applicationsarxiv-cs-ai
11 May 2026
Model Releases

MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development

DGX agent

arXiv:2603.24946v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance on automated software engineering tasks, yet existing benchmarks focus primarily on

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

On the Invariance and Generality of Neural Scaling Laws

DGX agent

arXiv:2605.07546v1 Announce Type: new Abstract: Neural scaling laws establish a predictable relationship between model performance and data or compute, offering crucial guidance for resource allocatio

model-releasesarxiv-cs-lg
11 May 2026
Safety

PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents

DGX agent

arXiv:2605.07039v1 Announce Type: new Abstract: Large language models have become drivers of evolutionary search, but most systems rely on a fixed, prompt-elicited policy to sample next candidates. Th

safetyarxiv-cs-lg
11 May 2026
Safety

PLOT: Progressive Localization via Optimal Transport in Neural Causal Abstraction

DGX agent

arXiv:2605.06979v1 Announce Type: cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning a high-level causal model with the low-level computation r

safetyarxiv-cs-ai
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ProcObject-10K: Benchmarking Object-Centric Procedural Understanding in Instructional Videos

DGX agent

arXiv:2512.03479v2 Announce Type: replace Abstract: Procedural activities are fundamentally driven by object state transitions, yet existing instructional video benchmarks remain action-centric and ca

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SatSurfGS: Generalizable 2D Gaussian Splatting for Sparse-View Satellite Surface Reconstruction

DGX agent

arXiv:2605.07181v1 Announce Type: new Abstract: Sparse-view satellite image surface reconstruction remains highly challenging, fundamentally because the reliability of multi-view matching under satell

model-releasesarxiv-cs-cv
11 May 2026
Local Ai

Self Driving Datasets: From 20 Million Papers to Nuanced Biomedical Knowledge at Scale

DGX agent

arXiv:2605.07022v1 Announce Type: new Abstract: Manually curated biomedical repositories -- spanning bioactivity, genomics, and chemistry -- are expensive to maintain, lag behind primary literature, a

local-aiarxiv-cs-lg
11 May 2026
Hardware

Sparser, Faster, Lighter Transformer Language Models

DGX agent

arXiv:2603.23198v2 Announce Type: replace-cross Abstract: Scaling autoregressive large language models (LLMs) has driven unprecedented progress but comes with vast computational costs. In this work, w

hardwarearxiv-cs-cl
11 May 2026
Model Releases

SREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios

DGX agent

arXiv:2605.07161v1 Announce Type: new Abstract: AI agents are increasingly used to diagnose and mitigate failures in production systems, known as agentic Site Reliability Engineering (SRE). Current SR

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos

DGX agent

arXiv:2605.07593v1 Announce Type: new Abstract: Real-world audio-visual understanding requires chaining evidence that is sparse, temporally dispersed, and split across the visual and auditory streams,

model-releasesarxiv-cs-cv
11 May 2026
Local Ai

Tracking Large-scale Shared Bikes with Inertial Motion Learning in GNSS Blocked Environments

DGX agent

arXiv:2605.07412v1 Announce Type: cross Abstract: Although Global Navigation Satellite Systems (GNSS) provide a general solution for bike tracking outdoors, there still exist complex riding environmen

local-aiarxiv-cs-ai
11 May 2026
Model Releases

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios

DGX agent

arXiv:2410.06355v3 Announce Type: replace-cross Abstract: This paper presents UNCOM, a novel hybrid framework for interpreting natural human commands in tabletop scenarios. The system integrates multi

model-releasesarxiv-cs-ai
11 May 2026
Applications

User eXperience Perception Insights Dataset (UXPID): Synthetic User Feedback from Public Industrial Forums

DGX agent

arXiv:2509.11777v2 Announce Type: replace Abstract: Customer feedback in industrial forums offers rich but underexplored insights into real-world product experience. Yet systematic analysis remains ch

applicationsarxiv-cs-cl
11 May 2026
Hardware

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

DGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

hardwarearxiv-cs-lg
11 May 2026
Model Releases

WiCER: Wiki-memory Compile, Evaluate, Refine Iterative Knowledge Compilation for LLM Wiki Systems

DGX agent

arXiv:2605.07068v1 Announce Type: cross Abstract: The LLM Wiki pattern, to compile and provide domain knowledge into a persistent artifact and serve it to LLMs via KV cache inference, promises context

model-releasesarxiv-cs-ai
11 May 2026
Applications

A Hybrid Method for Low-Resource Named Entity Recognition

DGX agent

arXiv:2605.04489v1 Announce Type: cross Abstract: Named Entity Recognition (NER) is a critical component of Natural Language Processing with diverse applications in information extraction and conversa

applicationsarxiv-cs-cl
7 May 2026
Local Ai

ARISE: A Repository-level Graph Representation and Toolset for Agentic Fault Localization and Program Repair

DGX agent

arXiv:2605.03117v1 Announce Type: cross Abstract: Repository-level fault localization (FL) and automated program repair (APR) require an agent to identify the relevant code units across files, follow

local-aiarxiv-cs-ai
7 May 2026
Model Releases

Benchmarking POS Tagging for the Tajik Language: A Comparative Study of Neural Architectures on the TajPersParallel Corpus

DGX agent

arXiv:2605.04576v1 Announce Type: new Abstract: This paper presents the first benchmark for the task of automatic part-of-speech (POS) tagging for the Tajik language. Despite the existence of multilin

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Cross-Model Consistency of Feature Importance in Electrospinning: Separating Robust from Model-Dependent Features

DGX agent

arXiv:2605.04905v1 Announce Type: new Abstract: Electrospinning is a highly sensitive fabrication process in which small variations in operating parameters can significantly influence fiber morphology

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

DiffCap-Bench: A Comprehensive, Challenging, Robust Benchmark for Image Difference Captioning

DGX agent

arXiv:2605.04503v1 Announce Type: new Abstract: Image Difference Captioning (IDC) generates natural language descriptions that precisely identify differences between two images, serving as a key bench

model-releasesarxiv-cs-cv
7 May 2026
Applications

Do Multimodal RAG Systems Leak Data? A Comprehensive Evaluation of Membership Inference and Image Caption Retrieval Attacks

DGX agent

arXiv:2601.17644v3 Announce Type: replace-cross Abstract: The growing adoption of multimodal Retrieval-Augmented Generation (mRAG) pipelines for vision-centric tasks (e.g., visual QA) introduces impor

applicationsarxiv-cs-ai
7 May 2026
Safety

Efficiently Aligning Language Models with Online Natural Language Feedback

DGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

safetyarxiv-cs-lg
7 May 2026
Model Releases

Fragile Knowledge, Robust Instruction-Following: The Width Pruning Dichotomy in Llama-3.2

DGX agent

arXiv:2512.22671v2 Announce Type: replace Abstract: Structured width pruning of GLU-MLP layers, guided by the Maximum Absolute Weight (MAW) criterion, reveals a systematic dichotomy in how reducing th

model-releasesarxiv-cs-cl
7 May 2026
Tutorials

From Diffusion to Rectified Flow: Rethinking Text-Based Segmentation

DGX agent

arXiv:2605.04590v1 Announce Type: new Abstract: Text-based image segmentation aims to delineate object boundaries within an image from text prompts, offering higher flexibility and broader application

tutorialsarxiv-cs-cv
7 May 2026
Model Releases

Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction

DGX agent

arXiv:2605.04770v1 Announce Type: new Abstract: While zero-shot appearance-based 3D gaze estimation offers significant cost-efficiency by directly mapping RGB images to gaze vectors, its reliability i

model-releasesarxiv-cs-cv
7 May 2026
Local Ai

HERCULES: Hardware-Efficient, Robust, Continual Learning Neural Architecture Search

DGX agent

arXiv:2605.04103v1 Announce Type: cross Abstract: Neural Architecture Search (NAS) has emerged as a powerful framework for automatically discovering neural architectures that balance accuracy and effi

local-aiarxiv-cs-cl
7 May 2026
Tutorials

HistoMet: A Pan-Cancer Deep Learning Framework for Prognostic Prediction of Metastatic Progression and Site Tropism from Primary Tumor Histopathology

DGX agent

arXiv:2602.07608v2 Announce Type: replace Abstract: Metastatic Progression remains the leading cause of cancer-related mortality, yet predicting whether a primary tumor will metastasize and where it w

tutorialsarxiv-cs-cv
7 May 2026
Model Releases

Imagery Dataset for Remaining Useful Life Estimation of Synthetic Fibre Ropes

DGX agent

arXiv:2605.04262v1 Announce Type: new Abstract: Remaining useful life (RUL) estimation of synthetic fibre ropes (SFRs) is critical for safe operation in offshore-crane, wind turbine installation, and

model-releasesarxiv-cs-cv
7 May 2026
Safety

Improving Bias Correction Standards by Quantifying its Effects on Treatment Outcomes

DGX agent

arXiv:2407.14861v3 Announce Type: replace-cross Abstract: With the growing access to administrative health databases, retrospective studies have become crucial evidence for medical treatments. Yet, no

safetyarxiv-cs-lg
7 May 2026
Safety

Investigating Trustworthiness of Nonparametric Deep Survival Models for Alzheimer's Disease Progression Analysis

DGX agent

arXiv:2605.04063v1 Announce Type: new Abstract: Alzheimer's Dementia (AD) is a progressive neurodegenerative disease marked by irreversible decline, making reliable modeling of its progression essenti

safetyarxiv-cs-lg
7 May 2026
Safety

LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey

DGX agent

arXiv:2505.00753v5 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing interest in building fully autonomous agents. However, fully autonomous LLM-bas

safetyarxiv-cs-cl
7 May 2026
Agents

Material Database Agent: A Multimodal Agentic Framework for Scientific Literature Mining

DGX agent

arXiv:2605.04278v1 Announce Type: new Abstract: Materials science workflows rely on structured and unstructured data from the vast body of available scientific literature. However, most of the experim

agentsarxiv-cs-cl
7 May 2026
Local Ai

MixINN: Accelerating Plant Breeding by Combining Mixed Models and Deep Learning for Interaction Prediction

DGX agent

arXiv:2605.04744v1 Announce Type: new Abstract: Plant breeding underpins global food security through incremental, accumulating improvements in crop yield, quality and sustainability, achieved via rep

local-aiarxiv-cs-lg
7 May 2026
Model Releases

MRI-Eval: A Tiered Benchmark for Evaluating LLM Performance on MRI Physics and GE Scanner Operations Knowledge

DGX agent

arXiv:2605.05175v1 Announce Type: cross Abstract: Background: Existing MRI LLM benchmarks rely mainly on review-book multiple-choice questions, where top proprietary models already score highly, limit

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Nsanku: Evaluating Zero-Shot Translation Performance of LLMs for Ghanaian Languages

DGX agent

arXiv:2605.04208v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated impressive multilingual capabilities for well-resourced languages, yet their performance on low-resource

model-releasesarxiv-cs-cl
7 May 2026
Agents

OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents

DGX agent

arXiv:2605.05185v1 Announce Type: new Abstract: Deep search has become a crucial capability for frontier multimodal agents, enabling models to solve complex questions through active search, evidence v

agentsarxiv-cs-cv
7 May 2026
Safety

OracleProto: A Reproducible Framework for Benchmarking LLM Native Forecasting via Knowledge Cutoff and Temporal Masking

DGX agent

arXiv:2605.03762v1 Announce Type: new Abstract: Large language models are moving from static text generators toward real-world decision-support systems, where forecasting is a composite capability tha

safetyarxiv-cs-ai
7 May 2026
Local Ai

Position: Embodied AI Requires a Privacy-Utility Trade-off

DGX agent

arXiv:2605.05017v1 Announce Type: cross Abstract: Embodied AI (EAI) systems are rapidly transitioning from simulations into real-world domestic and other sensitive environments. However, recent EAI so

local-aiarxiv-cs-ro
7 May 2026
Safety

Practical validation of synthetic pre-crash scenarios

DGX agent

arXiv:2605.04564v1 Announce Type: new Abstract: The representativeness of synthetic pre-crash scenarios is crucial for assessing the safety impact of Driving Automation Systems through virtual simulat

safetyarxiv-cs-ro
7 May 2026
Safety

Safety by Invariance, Liveness through Refinement: Heterogeneous Contract Framework for Co-Design of Layered Control

DGX agent

arXiv:2605.04222v1 Announce Type: cross Abstract: Real-world control systems must achieve long-horizon objectives (liveness) while respecting continuous-time safety constraints, a combination that mot

safetyarxiv-cs-ro
7 May 2026
Safety

Safety Must Precede the Deployment of Open-Ended AI

DGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

safetyarxiv-cs-ai
7 May 2026
Safety

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization

DGX agent

arXiv:2605.04700v1 Announce Type: cross Abstract: Jailbreak attacks on audio language models (ALMs) optimize audio perturbations to elicit unsafe generations, and they typically update the entire wave

safetyarxiv-cs-cl
7 May 2026
Model Releases

StoryAlign: Evaluating and Training Reward Models for Story Generation

DGX agent

arXiv:2605.04831v1 Announce Type: new Abstract: Story generation aims to automatically produce coherent, structured, and engaging narratives. Although large language models (LLMs) have significantly a

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

SWAN: Semantic Watermarking with Abstract Meaning Representation

DGX agent

arXiv:2605.04305v1 Announce Type: new Abstract: We introduce SWAN (Semantic Watermarking with Abstract Meaning Representation), a novel framework that embeds watermark signatures into the semantic str

model-releasesarxiv-cs-cl
7 May 2026
← Previous
1…443444445446447…462
Next →