AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
21 Apr 2026

SynthPID: P&ID digitization from Topology-Preserving Synthetic Data

Model ReleasesDGX agent

arXiv:2604.16513v1 Announce Type: new Abstract: Automating the digitization of Piping and Instrumentation Diagrams (P&IDs) into structured process graphs would unlock significant value in plant operat

The very first quadcopter prototype was the Breguet-Richet Gyroplane No. 1, created in 1907 by Breguet Aviation in France. The first consume…

Model ReleasesDGX agent

The very first quadcopter prototype was the Breguet-Richet Gyroplane No. 1, created in 1907 by Breguet Aviation in France. The first consumer quadcopter drone was the Parrot AR.Drone, released in 2010

Tight Clusters Make Specialized Experts

ResearchDGX agent

arXiv:2502.15315v3 Announce Type: replace Abstract: Sparse Mixture-of-Experts (MoE) architectures have emerged as a promising approach to decoupling model capacity from computational cost. At the core

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Time Series Forecasting as Reasoning: A Slow-Thinking Approach with Reinforced LLMs

SafetyDGX agent

arXiv:2506.10630v2 Announce Type: replace Abstract: To advance time series forecasting (TSF), various methods have been proposed to improve prediction accuracy, evolving from statistical techniques to

TowerDataset: A Heterogeneous Benchmark for Transmission Corridor Segmentation with a Global-Local Fusion Framework

Model ReleasesDGX agent

arXiv:2604.16848v1 Announce Type: new Abstract: Fine-grained semantic segmentation of transmission-corridor point clouds is fundamental for intelligent power-line inspection. However, current progress

Tree of Concepts: Interpretable Continual Learners in Non-Stationary Clinical Domains

ApplicationsDGX agent

arXiv:2604.17089v1 Announce Type: new Abstract: Continual learning aims to update models under distribution shift without forgetting, yet many high-stakes deployments, such as healthcare, also require

Unified Multimodal Brain Decoding via Cross-Subject Soft-ROI Fusion

Model ReleasesDGX agent

arXiv:2512.20249v3 Announce Type: replace-cross Abstract: Multimodal brain decoding aims to reconstruct semantic information that is consistent with visual stimuli from brain activity signals such as

Unveiling Deepfakes: A Frequency-Aware Triple Branch Network for Deepfake Detection

Model ReleasesDGX agent

arXiv:2604.17477v1 Announce Type: new Abstract: Advanced deepfake technologies are blurring the lines between real and fake, presenting both revolutionary opportunities and alarming threats. While it

VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Model ReleasesDGX agent

arXiv:2402.13243v2 Announce Type: replace Abstract: Learning a human-like driving policy from large-scale driving demonstrations is promising, but the uncertainty and non-deterministic nature of plann

Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation

Local AiDGX agent

arXiv:2604.17656v1 Announce Type: cross Abstract: Video-to-music (V2M) is the fundamental task of creating background music for an input video. Recent V2M models achieve audiovisual alignment by typic

VIDEOP2R: Video Understanding from Perception to Reasoning

SafetyDGX agent

arXiv:2511.11113v2 Announce Type: replace Abstract: Reinforcement fine-tuning (RFT), a two-stage framework consisting of supervised fine-tuning (SFT) and reinforcement learning (RL) has shown promisin

Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception

SafetyDGX agent

arXiv:2604.17475v1 Announce Type: cross Abstract: Small Vision-Language Models (SVLMs) are efficient task controllers but often suffer from visual brittleness and poor tool orchestration. They typical

We're open-sourcing FlashKDA — our high-performance CUTLASS-based implementation of Kimi Delta Attention kernels. Achieves 1.72×–2.22× prefi…

Model ReleasesDGX agent

We're open-sourcing FlashKDA — our high-performance CUTLASS-based implementation of Kimi Delta Attention kernels. Achieves 1.72×–2.22× prefill speedup over the flash-linear-attention baseline on H20,

Where is the Mind? Persona Vectors and LLM Individuation

ResearchDGX agent

arXiv:2604.17031v1 Announce Type: new Abstract: The individuation problem for large language models asks which entities associated with them, if any, should be identified as minds. We approach this pr

Who Gets the Kidney? Human-AI Alignment, Indecision, and Moral Values

SafetyDGX agent

arXiv:2506.00079v2 Announce Type: replace-cross Abstract: The rapid integration of Large Language Models (LLMs) in high-stakes decision-making -- such as allocating scarce resources like donor organs

20 Apr 2026

Attention Sinks Are Provably Necessary in Softmax Transformers: Evidence from Trigger-Conditional Tasks

ResearchDGX agent

arXiv:2603.11487v5 Announce Type: replace Abstract: Transformers often display an attention sink: probability mass concentrates on a fixed, content-agnostic position. Are sinks a byproduct of the opti

Beyond generating high-fidelity visuals, we wanted to test the limits of what Nano Banana Pro can do. We worked with design partners Porto R…

Model ReleasesDGX agent

Beyond generating high-fidelity visuals, we wanted to test the limits of what Nano Banana Pro can do. We worked with design partners Porto Rocha to build out a hypothetical brand called YOYOYO to see

BioHiCL: Hierarchical Multi-Label Contrastive Learning for Biomedical Retrieval with MeSH Labels

ResearchDGX agent

arXiv:2604.15591v1 Announce Type: cross Abstract: Effective biomedical information retrieval requires modeling domain semantics and hierarchical relationships among biomedical texts. Existing biomedic

CHOP: Chunkwise Context-Preserving Framework for RAG on Multi Documents

Model ReleasesDGX agent

arXiv:2604.15802v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems lose retrieval accuracy when similar documents coexist in the vector database, causing unnecessary informat

CodeMMR: Bridging Natural Language, Code, and Image for Unified Retrieval

Model ReleasesDGX agent

arXiv:2604.15663v1 Announce Type: cross Abstract: Code search, framed as information retrieval (IR), underpins modern software engineering and increasingly powers retrieval-augmented generation (RAG),

COMPASS: Benchmarking Constrained Optimization in LLM Agents

Model ReleasesDGX agent

arXiv:2510.07043v2 Announce Type: replace Abstract: Human decision-making often involves constrained optimization. As LLM agents are deployed to assist with real-world tasks like travel planning, shop

Creating and Evaluating Personas Using Generative AI: A Scoping Review of 81 Articles

ResearchDGX agent

arXiv:2504.04927v2 Announce Type: replace-cross Abstract: As generative AI (GenAI) is increasingly applied in persona development to represent real users, understanding the implications and limitation

DASB -- Discrete Audio and Speech Benchmark

Model ReleasesDGX agent

arXiv:2406.14294v3 Announce Type: replace-cross Abstract: Discrete audio tokens have recently gained considerable attention for their potential to bridge audio and language processing, enabling multim

DenTab: A Dataset for Table Recognition and Visual QA on Real-World Dental Estimates

Model ReleasesDGX agent

arXiv:2604.16099v1 Announce Type: new Abstract: Tables condense key transactional and administrative information into compact layouts, but practical extraction requires more than text recognition: sys

Designing Synthetic Discussion Generation Systems: A Case Study for Online Facilitation

ApplicationsDGX agent

arXiv:2503.16505v4 Announce Type: replace-cross Abstract: A critical challenge in social science research is the high cost associated with experiments involving human participants. We identify Synthet

Diffusion Autoencoder for Unsupervised Artifact Restoration in Handheld Fundus Images

TutorialsDGX agent

arXiv:2604.15723v1 Announce Type: cross Abstract: The advent of handheld fundus imaging devices has made ophthalmologic diagnosis and disease screening more accessible, efficient, and cost-effective.

Discovering quantum phenomena with Interpretable Machine Learning

TutorialsDGX agent

arXiv:2604.16015v1 Announce Type: cross Abstract: Interpretable machine learning techniques are becoming essential tools for extracting physical insights from complex quantum data. We build on recent

Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning

SafetyDGX agent

arXiv:2505.16176v2 Announce Type: replace Abstract: In mathematical reasoning, data selection strategies predominantly rely on static, externally defined metrics, which fail to adapt to the evolving c

Evaluating LLMs as Human Surrogates in Controlled Experiments

ResearchDGX agent

arXiv:2604.15329v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human responses in behavioral research, yet it remains unclear when LLM-generated data

Fed3D: Federated 3D Object Detection

Local AiDGX agent

arXiv:2604.15795v1 Announce Type: new Abstract: 3D object detection models trained in one server plays an important role in autonomous driving, robotics manipulation, and augmented reality scenarios.

FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation

Model ReleasesDGX agent

arXiv:2604.16298v1 Announce Type: new Abstract: UAV vision-language navigation (VLN) requires an agent to navigate complex 3D environments from an egocentric perspective while following ambiguous mult

Getting LLMs to simulate “true” randomness or generate diverse outputs is surprisingly difficult. We found a simple prompting trick that sol…

ResearchDGX agent

Getting LLMs to simulate “true” randomness or generate diverse outputs is surprisingly difficult. We found a simple prompting trick that solves this by having the model generate and manipulate a rando

Hero-Mamba: Mamba-based Dual Domain Learning for Underwater Image Enhancement

Model ReleasesDGX agent

arXiv:2604.16266v1 Announce Type: new Abstract: Underwater images often suffer from severe degradation, such as color distortion, low contrast, and blurred details, due to light absorption and scatter

Heterogeneous Sheaf Neural Networks

Model ReleasesDGX agent

arXiv:2409.08036v2 Announce Type: replace Abstract: Heterogeneous graphs, whose nodes and edges may belong to different types and feature spaces, arise in a wide variety of real-world domains such as

Intelligent Healthcare Imaging Platform: A VLM-Based Framework for Automated Medical Image Analysis and Clinical Report Generation

Model ReleasesDGX agent

arXiv:2509.13590v3 Announce Type: replace-cross Abstract: The rapid advancement of artificial intelligence (AI) in healthcare imaging has revolutionized diagnostic medicine and clinical decision-makin

IPQA: A Benchmark for Core Intent Identification in Personalized Question Answering

Model ReleasesDGX agent

arXiv:2510.23536v2 Announce Type: replace Abstract: Intent identification serves as the foundation for generating appropriate responses in personalized question answering (PQA). However, existing benc

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments

Model ReleasesDGX agent

arXiv:2604.15384v1 Announce Type: cross Abstract: We introduce LinuxArena, a control setting in which agents operate directly on live, multi-service production environments. LinuxArena contains 20 env

LLM Reasoning Is Latent, Not the Chain of Thought

ResearchDGX agent

arXiv:2604.15726v1 Announce Type: new Abstract: This position paper argues that large language model (LLM) reasoning should be studied as latent-state trajectory formation rather than as faithful surf

Natural gradient descent with momentum

Model ReleasesDGX agent

arXiv:2604.15554v1 Announce Type: cross Abstract: We consider the problem of approximating a function by an element of a nonlinear manifold which admits a differentiable parametrization, typical examp

Neurosymbolic Repo-level Code Localization

Model ReleasesDGX agent

arXiv:2604.16021v1 Announce Type: cross Abstract: Code localization is a cornerstone of autonomous software engineering. Recent advancements have achieved impressive performance on real-world issue be

Phase Transitions as the Breakdown of Statistical Indistinguishability

Model ReleasesDGX agent

arXiv:2604.15773v1 Announce Type: cross Abstract: We introduce a novel characterization of phase transitions based on hypothesis testing. In our formulation, a phase transition is defined as the break

Philosophy (among other things) grad here. I could write a whole essay about this video, and mostly the reactions to it. People are dunking …

Model ReleasesDGX agent

Philosophy (among other things) grad here. I could write a whole essay about this video, and mostly the reactions to it. People are dunking on her because of what she symbolises more than what she say

PIIBench: A Unified Multi-Source Benchmark Corpus for Personally Identifiable Information Detection

Model ReleasesDGX agent

arXiv:2604.15776v1 Announce Type: cross Abstract: We present PIIBench, a unified benchmark corpus for Personally Identifiable Information (PII) detection in natural language text. Existing resources f

Resource-efficient equivariant quantum convolutional neural networks

ResearchDGX agent

arXiv:2410.01252v2 Announce Type: replace-cross Abstract: Equivariant quantum neural networks (QNNs) are promising variational models that exploit symmetries to improve machine learning capabilities.

Robust Multispectral Semantic Segmentation under Missing or Full Modalities via Structured Latent Projection

SafetyDGX agent

arXiv:2604.15856v1 Announce Type: cross Abstract: Multimodal remote sensing data provide complementary information for semantic segmentation, but in real-world deployments, some modalities may be unav

SCHK-HTC: Sibling Contrastive Learning with Hierarchical Knowledge-Aware Prompt Tuning for Hierarchical Text Classification

Model ReleasesDGX agent

arXiv:2604.15998v1 Announce Type: new Abstract: Few-shot Hierarchical Text Classification (few-shot HTC) is a challenging task that involves mapping texts to a predefined tree-structured label hierarc

Seeing the imagined: a latent functional alignment in visual imagery decoding from fMRI data

Model ReleasesDGX agent

arXiv:2604.15374v1 Announce Type: cross Abstract: Recent progress in visual brain decoding from fMRI has been enabled by large-scale datasets such as the Natural Scenes Dataset (NSD) and powerful diff

Sequential KV Cache Compression via Probabilistic Language Tries: Beyond the Per-Vector Shannon Limit

ResearchDGX agent

arXiv:2604.15356v1 Announce Type: cross Abstract: Recent work on KV cache quantization, culminating in TurboQuant, has approached the Shannon entropy limit for per-vector compression of transformer ke

SIMMER: Cross-Modal Food Image--Recipe Retrieval via MLLM-Based Embedding

SafetyDGX agent

arXiv:2604.15628v1 Announce Type: cross Abstract: Cross-modal retrieval between food images and recipe texts is an important task with applications in nutritional management, dietary logging, and cook

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

Model ReleasesDGX agent

arXiv:2604.15735v1 Announce Type: cross Abstract: Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-dr

Softpick: No Attention Sink, No Massive Activations with Rectified Softmax

Model ReleasesDGX agent

arXiv:2504.20966v4 Announce Type: replace Abstract: We introduce softpick, a rectified, not sum-to-one, drop-in replacement for softmax in transformer attention mechanisms that eliminates attention si

SSFT: A Lightweight Spectral-Spatial Fusion Transformer for Generic Hyperspectral Classification

Model ReleasesDGX agent

arXiv:2604.15828v1 Announce Type: new Abstract: Hyperspectral imaging enables fine-grained recognition of materials by capturing rich spectral signatures, but learning robust classifiers is challengin

Target-Oriented Pretraining Data Selection via Neuron-Activated Graph

ResearchDGX agent

arXiv:2604.15706v1 Announce Type: new Abstract: Everyday tasks come with a target, and pretraining models around this target is what turns them into experts. In this paper, we study target-oriented la

The Illusion of Equivalence: Systematic FP16 Divergence in KV-Cached Autoregressive Inference

Model ReleasesDGX agent

arXiv:2604.15409v1 Announce Type: cross Abstract: KV caching is a ubiquitous optimization in autoregressive transformer inference, long presumed to be numerically equivalent to cache-free computation.

The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement

Model ReleasesDGX agent

arXiv:2604.16116v1 Announce Type: cross Abstract: We extracted the scholarly reasoning systems of two internationally prominent humanities and social science scholars from their published corpora alon

TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.15950v1 Announce Type: new Abstract: Pancreatic ductal adenocarcinoma (PDAC) segmentation on contrast-enhanced CT is inherently ambiguous: inter-rater disagreement among experts reflects ge

UA-Net: Uncertainty-Aware Network for TRISO Image Semantic Segmentation

ResearchDGX agent

arXiv:2604.15542v1 Announce Type: new Abstract: Tristructural isotropic (TRISO)-coated particle fuels undergo dimensional changes and chemical reactions during high-temperature neutron irradiation. Po

Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation

ResearchDGX agent

arXiv:2511.02626v3 Announce Type: replace Abstract: Prior works have shown that fine-tuning on new knowledge can induce factual hallucinations in large language models (LLMs), leading to incorrect out

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

AgentsDGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

VeRVE: Versatile Retrieval for Videos via Unified Embeddings

Local AiDGX agent

arXiv:2601.12193v3 Announce Type: replace Abstract: Modern video retrieval systems are expected to handle diverse tasks ranging from corpus-level retrieval, fine-grained moment localization to flexibl

← Previous
1…493494495496497…1071
Next →