AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Model Releases

Can Current Agents Close the Discovery-to-Application Gap? A Case Study in Minecraft

DGX agent

arXiv:2604.24697v1 Announce Type: new Abstract: Discovering causal regularities and applying them to build functional systems--the discovery-to-application loop--is a hallmark of general intelligence,

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters

DGX agent

arXiv:2604.24710v1 Announce Type: new Abstract: Objective. Clinical AI documentation systems require evaluation methodologies that are clinically valid, economically viable, and sensitive to iterative

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Cataract-LMM Large-Scale Multi-Source Multi-Task Benchmark for Deep Learning in Surgical Video Analysis

DGX agent

arXiv:2510.16371v2 Announce Type: replace-cross Abstract: The development of computer-assisted surgery systems relies on large-scale, annotated datasets. Existing cataract surgery resources lack the d

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads

DGX agent

arXiv:2604.23009v1 Announce Type: new Abstract: Job Skill Named Entity Recognition (JobSkillNER) aims to automatically extract key skill information from large-scale job posting data, which is importa

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

ClawTrace: Cost-Aware Tracing for LLM Agent Skill Distillation

DGX agent

arXiv:2604.23853v1 Announce Type: new Abstract: Skill-distillation pipelines learn reusable rules from LLM agent trajectories, but they lack a key signal: how much each step costs. Without per-step co

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Comparative Insights on Adversarial Machine Learning from Industry and Academia: A User-Study Approach

DGX agent

arXiv:2602.04753v2 Announce Type: replace-cross Abstract: An exponential growth of Machine Learning and its Generative AI applications brings with it significant security challenges, often referred to

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

Complexity of Linear Regions in Self-supervised Deep ReLU Networks

DGX agent

arXiv:2604.24393v1 Announce Type: cross Abstract: There has been growing interest in studying the complexity of Rectified Linear Unit (ReLU) based activation networks. Recent work investigates the evo

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Computational Design and Co-Robotic Fabrication for Material Reuse in Architecture

DGX agent

arXiv:2604.24648v1 Announce Type: new Abstract: Climate change and resource depletion demand a shift from the dominant linear 'take-make-use-dispose' paradigm of construction toward circular, low-wast

applicationsarxiv-cs-ro
28 Apr 2026
Model Releases

DGHMesh: A Large-scale Dual-radar mmWave Dataset and Generalization-focused Benchmark for Human Mesh Reconstruction

DGX agent

arXiv:2604.22827v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar has shown great potential for contactless, privacy-preserving, and robust human sensing, yet existing mmWave-based human

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations?

DGX agent

arXiv:2604.23688v1 Announce Type: new Abstract: Proactive defense methods protect portrait images from unauthorized editing or talking face generation (TFG) by introducing pixel-level protective pertu

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

DyABD: The Abdominal Muscle Segmentation in Dynamic MRI Benchmark

DGX agent

arXiv:2604.23187v1 Announce Type: cross Abstract: This work introduces DyABD, a novel and complex benchmark dataset of dynamic abdominal MRIs from patients with abdominal hernias and associated high q

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

DGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models

DGX agent

arXiv:2502.04424v4 Announce Type: replace-cross Abstract: With the integration of multimodal large language models (MLLMs) into robotic systems and AI applications, embedding emotional intelligence (E

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs

DGX agent

arXiv:2604.23348v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and generation, and are increasingly used in

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Evaluating the Search Agent in a Parallel World

DGX agent

arXiv:2603.04751v2 Announce Type: replace Abstract: Integrating web search tools has significantly extended the capability of LLMs to address open-world, real-time, and long-tail problems. However, ev

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Exploring the Secondary Risks of Large Language Models

DGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

FAIR_XAI: Improving Multimodal Foundation Model Fairness via Explainability for Wellbeing Assessment

DGX agent

arXiv:2604.23786v1 Announce Type: new Abstract: In recent years, the integration of multimodal machine learning in wellbeing assessment has offered transformative potential for monitoring mental healt

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

FastAT Benchmark: A Comprehensive Framework for Fair Evaluation of Fast Adversarial Training Methods

DGX agent

arXiv:2604.22853v1 Announce Type: new Abstract: Fast Adversarial Training (FastAT) seeks to achieve adversarial robustness at a fraction of the computational cost incurred by standard multi-step metho

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification

DGX agent

arXiv:2604.23588v1 Announce Type: new Abstract: Financial AI systems must produce answers grounded in specific regulatory filings, yet current LLMs fabricate metrics, invent citations, and miscalculat

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Forecasting Commencing Enrolments Under Data Sparsity: A Zero-Shot Time Series Foundation Models Framework for Higher Education Planning

DGX agent

arXiv:2602.12120v3 Announce Type: replace Abstract: Effective resource allocation in higher education depends on reliable enrolment forecasts, yet institutional planners frequently face data series di

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

From Stateless Queries to Autonomous Actions: A Layered Security Framework for Agentic AI Systems

DGX agent

arXiv:2604.23338v1 Announce Type: cross Abstract: Agentic AI systems face security challenges that stateless large language models do not. They plan across extended horizons, maintain persistent memor

safetyarxiv-cs-lg
28 Apr 2026
Model Releases

Game-Time: Evaluating Temporal Dynamics in Spoken Language Models

DGX agent

arXiv:2509.26388v3 Announce Type: replace-cross Abstract: Conversational Spoken Language Models (SLMs) are emerging as a promising paradigm for real-time speech interaction. However, their capacity of

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

'If You're Very Clever, No One Knows You've Used It': The Social Dynamics of Developing Generative AI Literacy in the Workplace

DGX agent

arXiv:2602.01386v2 Announce Type: replace-cross Abstract: Generative AI (GenAI) tools are rapidly transforming knowledge work, making AI literacy a critical priority for organizations. However, resear

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

DGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

model-releasesarxiv-cs-ai
28 Apr 2026
Hardware

Latent Inter-Frame Pruning: A Training-Free Method Bridging Traditional Video Compression and Modern Diffusion Transformers for Efficient Generation

DGX agent

arXiv:2604.23858v1 Announce Type: new Abstract: Video generation, while capable of generating realistic videos, is computationally expensive and slow, prohibiting real-time applications. In this paper

hardwarearxiv-cs-cv
28 Apr 2026
Safety

Learning from Imperfect Text Guidance: Robust Long-Tail Visual Recognition with High-Noise Label

DGX agent

arXiv:2604.23125v1 Announce Type: new Abstract: Real-world data often exhibit long-tailed distributions with numerous noisy labels, substantially degrading the performance of deep models. While prior

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain

DGX agent

arXiv:2509.10546v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in finance, where unsafe behavior can lead to serious regulatory risks. However, most r

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation

DGX agent

arXiv:2604.23604v1 Announce Type: new Abstract: Understanding the surrounding environment is fundamental in autonomous driving and robotic perception. Distinguishing between known classes and previous

agentsarxiv-cs-cv
28 Apr 2026
Model Releases

LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews

DGX agent

arXiv:2511.12635v2 Announce Type: replace-cross Abstract: Context: Large language models (LLMs) are increasingly used to screen literature for systematic reviews (SRs), but the standard confusion-matr

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

OLaPh: Optimal Language Phonemizer

DGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

DGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving

DGX agent

arXiv:2604.23712v1 Announce Type: cross Abstract: Recent advances in formal theorem proving have focused on Olympiad-level mathematics, leaving undergraduate domains largely unexplored. Optimization,

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Orthogonal Representation Learning for Estimating Causal Quantities

DGX agent

arXiv:2502.04274v4 Announce Type: replace Abstract: End-to-end representation learning has become a powerful tool for estimating causal quantities from high-dimensional observational data, but its eff

safetyarxiv-cs-lg
28 Apr 2026
Safety

Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing

DGX agent

arXiv:2604.23412v1 Announce Type: new Abstract: While annotated corpora are crucial in the field of natural language processing (NLP), those containing copyrighted material are difficult to exchange a

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

DGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Probing Visual Planning in Image Editing Models

DGX agent

arXiv:2604.22868v1 Announce Type: cross Abstract: Visual planning represents a crucial facet of human intelligence, especially in tasks that require complex spatial reasoning and navigation. Yet, in m

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering

DGX agent

arXiv:2604.24052v1 Announce Type: cross Abstract: Video-to-text summarization remains underexplored in terms of comprehensive evaluation methods. Traditional n-gram overlap-based metrics and recent la

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

DGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization

DGX agent

arXiv:2408.00923v2 Announce Type: replace-cross Abstract: This paper explores a novel paradigm in low-bit (i.e. 4-bits or lower) quantization, differing from existing state-of-the-art methods, by fram

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Scalable Explainability-as-a-Service (XaaS) for Edge AI Systems

DGX agent

arXiv:2602.04120v2 Announce Type: replace-cross Abstract: Though Explainable AI (XAI) has made significant advancements, its inclusion in edge and IoT systems is typically ad-hoc and inefficient. Most

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services

DGX agent

arXiv:2604.24023v1 Announce Type: new Abstract: Recent image generation and editing models demonstrate robust adherence to instructions and high visual quality on academic benchmarks. However, their p

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning

DGX agent

arXiv:2604.23747v1 Announce Type: cross Abstract: Recent mixed-policy optimization methods for LLM reasoning that interleave or blend supervised and reinforcement learning signals report improvements

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning

DGX agent

arXiv:2506.05425v2 Announce Type: replace-cross Abstract: Understanding social interaction, which encompasses perceiving numerous and subtle multimodal cues, inferring unobservable mental states and r

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Skill Retrieval Augmentation for Agentic AI

DGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing

DGX agent

arXiv:2604.23392v1 Announce Type: new Abstract: Refereeing is vital in sports, where fair, accurate, and explainable decisions are fundamental. While intelligent assistant technologies are being widel

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SolarFCD: A Large-Scale Dataset and Benchmark for Solar Fault Classification in Photovoltaic Systems

DGX agent

arXiv:2604.23662v1 Announce Type: new Abstract: The increasing global deployment of solar photovoltaic (PV) systems needs robust, scalable, and automated inspection technologies capable of detecting a

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…448449450451452…462
Next →