AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Model Releases

CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild

DGX agent

arXiv:2604.08287v1 Announce Type: new Abstract: Discovering camouflaged objects is a challenging task in computer vision due to the high similarity between camouflaged objects and their surroundings.

model-releasesarxiv-cs-cv
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Can Vision Language Models Judge Action Quality? An Empirical Evaluation

DGX agent

arXiv:2604.08294v1 Announce Type: cross Abstract: Action Quality Assessment (AQA) has broad applications in physical therapy, sports coaching, and competitive judging. Although Vision Language Models

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild

DGX agent

arXiv:2604.07354v1 Announce Type: new Abstract: The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes do

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Data Leakage in Automotive Perception: Practitioners' Insights

DGX agent

arXiv:2604.06899v1 Announce Type: cross Abstract: Data leakage is the inadvertent transfer of information between training and evaluation datasets that poses a subtle, yet critical, risk to the reliab

safetyarxiv-cs-lg
10 Apr 2026
Agents

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather

DGX agent

arXiv:2604.08074v1 Announce Type: new Abstract: Reliable and weather-robust perception systems are essential for safe autonomous driving and typically employ multi-modal sensor configurations to achie

agentsarxiv-cs-cv
10 Apr 2026
Safety

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

DGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Face-D(^2)CL: Multi-Domain Synergistic Representation with Dual Continual Learning for Facial DeepFake Detection

DGX agent

arXiv:2604.08159v1 Announce Type: new Abstract: The rapid advancement of facial forgery techniques poses severe threats to public trust and information security, making facial DeepFake detection a cri

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

DGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

DGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

DGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

DGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

safetyarxiv-cs-cl
10 Apr 2026
Safety

Front-End Ethics for Sensor-Fused Health Conversational Agents: An Ethical Design Space for Biometrics

DGX agent

arXiv:2604.06203v1 Announce Type: cross Abstract: The integration of continuous data from built-in sensors and Large Language Models (LLMs) has fueled a surge of 'Sensor-Fused LLM agents' for personal

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

DGX agent

arXiv:2604.07429v1 Announce Type: new Abstract: Towards an embodied generalist for real-world interaction, Multimodal Large Language Model (MLLM) agents still suffer from challenging latency, sparse f

model-releasesarxiv-cs-cv
10 Apr 2026
Applications

GaussiAnimate: Reconstruct and Rig Animatable Categories with Level of Dynamics

DGX agent

arXiv:2604.08547v1 Announce Type: new Abstract: Free-form bones, that conform closely to the surface, can effectively capture non-rigid deformations, but lack a kinematic structure necessary for intui

applicationsarxiv-cs-cv
10 Apr 2026
Applications

Gaze to Insight: A Scalable AI Approach for Detecting Gaze Behaviours in Face-to-Face Collaborative Learning

DGX agent

arXiv:2604.03317v2 Announce Type: replace Abstract: Previous studies have illustrated the potential of analysing gaze behaviours in collaborative learning to provide educationally meaningful informati

applicationsarxiv-cs-cv
10 Apr 2026
Hardware

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

DGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

hardwarearxiv-cs-cv
10 Apr 2026
Applications

Horticultural Temporal Fruit Monitoring via 3D Instance Segmentation and Re-Identification using Colored Point Clouds

DGX agent

arXiv:2411.07799v3 Announce Type: replace Abstract: Accurate and consistent fruit monitoring over time is a key step toward automated agricultural production systems. However, this task is inherently

applicationsarxiv-cs-cv
10 Apr 2026
Model Releases

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

DGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

DGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

model-releasesarxiv-cs-ai
10 Apr 2026
Applications

LLM-Based Data Generation and Clinical Skills Evaluation for Low-Resource French OSCEs

DGX agent

arXiv:2604.08126v1 Announce Type: new Abstract: Objective Structured Clinical Examinations (OSCEs) are the standard method for assessing medical students' clinical and communication skills through str

applicationsarxiv-cs-cl
10 Apr 2026
Safety

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

DGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

DGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

LUMINA: Foundation Models for Topology Transferable ACOPF

DGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Matrix Profile for Anomaly Detection on Multidimensional Time Series

DGX agent

arXiv:2409.09298v2 Announce Type: replace-cross Abstract: The Matrix Profile (MP), a versatile tool for time series data mining, has been shown effective in time series anomaly detection (TSAD). This

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

DGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

agentsarxiv-cs-lg
10 Apr 2026
Safety

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

DGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

safetyarxiv-cs-cl
10 Apr 2026
Applications

Mining Electronic Health Records to Investigate Effectiveness of Ensemble Deep Clustering

DGX agent

arXiv:2604.07085v1 Announce Type: new Abstract: In electronic health records (EHRs), clustering patients and distinguishing disease subtypes are key tasks to elucidate pathophysiology and aid clinical

applicationsarxiv-cs-lg
10 Apr 2026
Model Releases

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

DGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

model-releasesarxiv-cs-cv
10 Apr 2026
Agents

MolmoWeb: Open Visual Web Agent and Open Data for the Open Web

DGX agent

arXiv:2604.08516v1 Announce Type: new Abstract: Web agents--autonomous systems that navigate and execute tasks on the web on behalf of users--have the potential to transform how people interact with t

agentsarxiv-cs-cv
10 Apr 2026
Agents

ORACLE-SWE: Quantifying the Contribution of Oracle Information Signals on SWE Agents

DGX agent

arXiv:2604.07789v1 Announce Type: cross Abstract: Recent advances in language model (LM) agents have significantly improved automated software engineering (SWE). Prior work has proposed various agenti

agentsarxiv-cs-cl
10 Apr 2026
Agents

ParkSense: Where Should a Delivery Driver Park? Leveraging Idle AV Compute and Vision-Language Models

DGX agent

arXiv:2604.07912v1 Announce Type: new Abstract: Finding parking consumes a disproportionate share of food delivery time, yet no system addresses precise parking-spot selection relative to merchant ent

agentsarxiv-cs-cv
10 Apr 2026
Safety

Personalizing Text-to-Image Generation to Individual Taste

DGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation

DGX agent

arXiv:2512.23994v3 Announce Type: replace-cross Abstract: Text-to-audio-video (T2AV) generation is central to applications such as filmmaking and world modeling. However, current models often fail to

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

DGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks

DGX agent

arXiv:2511.19474v4 Announce Type: replace-cross Abstract: Automatically detecting abnormal events in videos is crucial for modern autonomous systems, yet existing Video Anomaly Detection (VAD) benchma

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Prompt reinforcing for long-term planning of large language models

DGX agent

arXiv:2510.05921v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable success in a wide range of natural language processing tasks and can be adapted through prompt

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Quantitative Estimation of Target Task Performance from Unsupervised Pretext Task in Semi/Self-Supervised Learning

DGX agent

arXiv:2508.07299v2 Announce Type: replace-cross Abstract: The effectiveness of unlabeled data in Semi/Self-Supervised Learning (SSL) depends on appropriate assumptions for specific scenarios, thereby

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

DGX agent

arXiv:2604.07990v1 Announce Type: new Abstract: The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both seman

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

DGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving

DGX agent

arXiv:2604.08008v1 Announce Type: new Abstract: Retrieving rare and safety-critical driving scenarios from large-scale datasets is essential for building robust autonomous driving (AD) systems. As dat

model-releasesarxiv-cs-cv
10 Apr 2026
Applications

SMFD-UNet: Semantic Face Mask Is The Only Thing You Need To Deblur Faces

DGX agent

arXiv:2604.07477v1 Announce Type: new Abstract: For applications including facial identification, forensic analysis, photographic improvement, and medical imaging diagnostics, facial image deblurring

applicationsarxiv-cs-cv
10 Apr 2026
Tutorials

Steering the Verifiability of Multimodal AI Hallucinations

DGX agent

arXiv:2604.06714v1 Announce Type: new Abstract: AI applications driven by multimodal large language models (MLLMs) are prone to hallucinations and pose considerable risks to human users. Crucially, su

tutorialsarxiv-cs-ai
10 Apr 2026
Model Releases

Tabular GANs for uneven distribution

DGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

model-releasesarxiv-cs-cv
10 Apr 2026
Tutorials

TEC: A Collection of Human Trial-and-error Trajectories for Problem Solving

DGX agent

arXiv:2604.06734v2 Announce Type: replace Abstract: Trial-and-error is a fundamental strategy for humans to solve complex problems and a necessary capability for Artificial Intelligence (AI) systems o

tutorialsarxiv-cs-cl
10 Apr 2026
Agents

The Art of Building Verifiers for Computer Use Agents

DGX agent

arXiv:2604.06240v1 Announce Type: cross Abstract: Verifying the success of computer use agent (CUA) trajectories is a critical challenge: without reliable verification, neither evaluation nor training

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

The ATOM Report: Measuring the Open Language Model Ecosystem

DGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation

DGX agent

arXiv:2604.06205v1 Announce Type: cross Abstract: The growth of online platforms and user content requires strong content moderation systems that can handle complex inputs from various media types. Wh

safetyarxiv-cs-ai
10 Apr 2026
← Previous
1…459460461462
Next →