AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

PBE-UNet: A light weight Progressive Boundary-Enhanced U-Net with Scale-Aware Aggregation for Ultrasound Image Segmentation

DGX agent

arXiv:2604.13791v1 Announce Type: new Abstract: Accurate lesion segmentation in ultrasound images is essential for preventive screening and clinical diagnosis, yet remains challenging due to low contr

model-releasesarxiv-cs-cv
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PersonaVLM: Long-Term Personalized Multimodal LLMs

DGX agent

arXiv:2604.13074v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) serve as daily assistants for millions. However, their ability to generate responses aligned with individual pr

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization

DGX agent

arXiv:2508.00222v5 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has significantly advanced the complex reasoning abilities of Large Language Models (LLMs

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

Seedance 2.0: Advancing Video Generation for World Complexity

DGX agent

arXiv:2604.14148v1 Announce Type: new Abstract: Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecesso

model-releasesarxiv-cs-cv
16 Apr 2026
Research

SiLVR: A Simple Language-based Video Reasoning Framework

DGX agent

arXiv:2505.24869v3 Announce Type: replace Abstract: Recent advances in test-time optimization have led to remarkable reasoning capabilities in Large Language Models (LLMs), enabling them to solve high

researcharxiv-cs-cv
16 Apr 2026
Model Releases

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs

DGX agent

arXiv:2604.13710v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exhibit strong reasoning and world knowledge, yet adapting them for retrieval remains challenging. Existing app

model-releasesarxiv-cs-cv
16 Apr 2026
Research

The Signal is in the Steps: Local Scoring for Reasoning Data Selection

DGX agent

arXiv:2510.03988v2 Announce Type: replace Abstract: Distilling long-form reasoning from teacher models into smaller students requires selecting which candidate solutions to train on. Recent work argue

researcharxiv-cs-lg
16 Apr 2026
Tutorials

Tokenizing Semantic Segmentation with Run Length Encoding

DGX agent

arXiv:2602.21627v3 Announce Type: replace Abstract: This paper presents a new unified approach to semantic segmentation in both images and videos by using language modeling to output the masks as sequ

tutorialsarxiv-cs-cv
16 Apr 2026
Model Releases

Two-Stage Regularization-Based Structured Pruning for LLMs

DGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

model-releasesarxiv-cs-cl
16 Apr 2026
Local Ai

Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization

DGX agent

arXiv:2604.13197v1 Announce Type: new Abstract: Process reward models (PRMs) provide fine-grained reward signals along the reasoning process, but training reliable PRMs often requires step annotations

local-aiarxiv-cs-cl
16 Apr 2026
Model Releases

WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain

DGX agent

arXiv:2604.13055v1 Announce Type: new Abstract: Today's evolving labor markets rely increasingly on recommender systems for hiring, talent management, and workforce analytics, with natural language pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

A Sanity Check on Composed Image Retrieval

DGX agent

arXiv:2604.12904v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image, and a relative caption that specifies the

model-releasesarxiv-cs-cv
15 Apr 2026
Research

AdaMCoT: Rethinking Cross-Lingual Factual Reasoning through Adaptive Multilingual Chain-of-Thought

DGX agent

arXiv:2501.16154v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown impressive multilingual capabilities through pretraining on diverse corpora. Although these models sho

researcharxiv-cs-ai
15 Apr 2026
Model Releases

AISafetyBenchExplorer: A Metric-Aware Catalogue of AI Safety Benchmarks Reveals Fragmented Measurement and Weak Benchmark Governance

DGX agent

arXiv:2604.12875v1 Announce Type: new Abstract: The rapid expansion of large language model (LLM) safety evaluation has produced a substantial benchmark ecosystem, but not a correspondingly coherent m

model-releasesarxiv-cs-ai
15 Apr 2026
Agents

CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades

DGX agent

arXiv:2604.12262v1 Announce Type: cross Abstract: Cascaded LLM systems coordinate models of varying sizes with human experts to balance accuracy, cost, and abstention under uncertainty. However, singl

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

CoD-Lite: Real-Time Diffusion-Based Generative Image Compression

DGX agent

arXiv:2604.12525v1 Announce Type: new Abstract: Recent advanced diffusion methods typically derive strong generative priors by scaling diffusion transformers. However, scaling fails to generalize when

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

CoDe-R: Refining Decompiler Output with LLMs via Rationale Guidance and Adaptive Inference

DGX agent

arXiv:2604.12913v1 Announce Type: cross Abstract: Binary decompilation is a critical reverse engineering task aimed at reconstructing high-level source code from stripped executables. Although Large L

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation

DGX agent

arXiv:2604.12268v1 Announce Type: cross Abstract: Large language models (LLMs) can generate code from natural language, but the extent to which they capture intended program behavior remains unclear.

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

FAST-DIPS: Adjoint-Free Analytic Steps and Hard-Constrained Likelihood Correction for Diffusion-Prior Inverse Problems

DGX agent

arXiv:2603.01591v2 Announce Type: replace-cross Abstract: Training-free diffusion priors enable inverse-problem solvers without retraining, but for nonlinear forward operators data consistency often r

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

FMASH: Advancing Traditional Chinese Medicine Formula Recommendation with Efficient Fusion of Multiscale Associations of Symptoms and Herbs

DGX agent

arXiv:2503.05167v3 Announce Type: replace Abstract: Traditional Chinese medicine (TCM) exhibits remarkable therapeutic efficacy in healthcare through patient-specific formulas. However, current AI-bas

local-aiarxiv-cs-lg
15 Apr 2026
Model Releases

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

DGX agent

arXiv:2604.12666v1 Announce Type: cross Abstract: Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Fundus Image-based Glaucoma Screening via Retinal Knowledge-Oriented Dynamic Multi-Level Feature Integration

DGX agent

arXiv:2604.12351v1 Announce Type: new Abstract: Automated diagnosis based on color fundus photography is essential for large-scale glaucoma screening. However, existing deep learning models are typica

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

DGX agent

arXiv:2604.12630v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Rec

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

GTCN-G: A Residual Graph-Temporal Fusion Network for Imbalanced Intrusion Detection

DGX agent

arXiv:2510.07285v3 Announce Type: replace-cross Abstract: The escalating complexity of network threats and the inherent class imbalance in traffic data present formidable challenges for modern Intrusi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

IDEA: An Interpretable and Editable Decision-Making Framework for LLMs via Verbal-to-Numeric Calibration

DGX agent

arXiv:2604.12573v1 Announce Type: new Abstract: Large Language Models are increasingly deployed for decision-making, yet their adoption in high-stakes domains remains limited by miscalibrated probabil

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study

DGX agent

arXiv:2604.12337v1 Announce Type: new Abstract: Letters of recommendation (LoRs) can carry patterns of implicitly gendered language that can inadvertently influence downstream decisions, e.g. in hirin

model-releasesarxiv-cs-lg
15 Apr 2026
Research

IMU: Influence-guided Machine Unlearning

DGX agent

arXiv:2508.01620v3 Announce Type: replace-cross Abstract: Machine Unlearning (MU) aims to selectively erase the influence of specific data points from pretrained models. However, most existing MU meth

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Invariant Features for Global Crop Type Classification

DGX agent

arXiv:2509.03497v3 Announce Type: replace Abstract: Accurate global crop type mapping supports agricultural monitoring and food security, yet remains limited by the scarcity of labeled data in many re

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Inverse Design of Inorganic Compounds with Generative AI

DGX agent

arXiv:2604.11827v1 Announce Type: cross Abstract: Machine learning is revolutionizing chemistry. Beyond the value of predictive models accelerating virtual screening, generative AI aims at enabling in

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidance

DGX agent

arXiv:2604.12627v1 Announce Type: new Abstract: RLVR improves reasoning in large language models, but its effectiveness is often limited by severe reward sparsity on hard problems. Recent hint-based R

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Leveraging Weighted Syntactic and Semantic Context Assessment Summary (wSSAS) Towards Text Categorization Using LLMs

DGX agent

arXiv:2604.12049v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) for reliable, enterprise-grade analytics such as text categorization is often hindered by the stochastic natur

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

LLM-Guided Prompt Evolution for Password Guessing

DGX agent

arXiv:2604.12601v1 Announce Type: cross Abstract: Passwords still remain a dominant authentication method, yet their security is routinely subverted by predictable user choices and large-scale credent

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

MedConcept: Unsupervised Concept Discovery for Interpretability in Medical VLMs

DGX agent

arXiv:2604.11868v1 Announce Type: new Abstract: While medical Vision-Language models (VLMs) achieve strong performance on tasks such as tumor or organ segmentation and diagnosis prediction, their opaq

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Mema: Memory-Augmented Adapter for Enhanced Vision-Language Understanding

DGX agent

arXiv:2603.00655v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance by aligning pretrained visual representations with the linguistic know

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Orthogonal Subspace Projection for Continual Machine Unlearning via SVD-Based LoRA

DGX agent

arXiv:2604.12526v1 Announce Type: cross Abstract: Continual machine unlearning aims to remove the influence of data that should no longer be retained, while preserving the usefulness of the model on e

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Parametric Interpolation of Dynamic Mode Decomposition for Predicting Nonlinear Systems

DGX agent

arXiv:2604.12103v1 Announce Type: cross Abstract: We present parameter-interpolated dynamic mode decomposition (piDMD), a parametric reduced-order modeling framework that embeds known parameter-affine

model-releasesarxiv-cs-lg
15 Apr 2026
Hardware

PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving

DGX agent

arXiv:2604.12171v1 Announce Type: cross Abstract: Pipeline parallelism (PP) is widely used to partition layers of large language models (LLMs) across GPUs, enabling scalable inference for large models

hardwarearxiv-cs-lg
15 Apr 2026
Model Releases

ReflectCAP: Detailed Image Captioning with Reflective Memory

DGX agent

arXiv:2604.12357v1 Announce Type: new Abstract: Detailed image captioning demands both factual grounding and fine-grained coverage, yet existing methods have struggled to achieve them simultaneously.

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Should There be a Teacher In-the-Loop? A Study of Generative AI Personalized Tasks Middle School

DGX agent

arXiv:2602.15876v1 Announce Type: cross Abstract: Adapting instruction to the fine-grained needs of individual students is a powerful application of recent advances in large language models. These gen

researcharxiv-cs-ai
15 Apr 2026
Model Releases

SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks

DGX agent

arXiv:2506.14512v4 Announce Type: replace Abstract: Large Language Models (LLMs) have undergone rapid progress, largely attributed to reinforcement learning on complex reasoning tasks. In contrast, wh

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

Spatial-Spectral Adaptive Fidelity and Noise Prior Reduction Guided Hyperspectral Image Denoising

DGX agent

arXiv:2604.12600v1 Announce Type: new Abstract: The core challenge of hyperspectral image denoising is striking the right balance between data fidelity and noise prior modeling. Most existing methods

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

The Verification Tax: Fundamental Limits of AI Auditing in the Rare-Error Regime

DGX agent

arXiv:2604.12951v1 Announce Type: new Abstract: The most cited calibration result in deep learning -- post-temperature-scaling ECE of 0.012 on CIFAR-100 (Guo et al., 2017) -- is below the statistical

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Thought-Retriever: Don't Just Retrieve Raw Data, Retrieve Thoughts for Memory-Augmented Agentic Systems

DGX agent

arXiv:2604.12231v1 Announce Type: new Abstract: Large language models (LLMs) have transformed AI research thanks to their powerful internal capabilities and knowledge. However, existing LLMs still fai

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Unlocking the Potential of Grounding DINO in Videos: Parameter-Efficient Adaptation for Limited-Data Spatial-Temporal Localization

DGX agent

arXiv:2604.12346v1 Announce Type: new Abstract: Spatio-temporal video grounding (STVG) aims to localize queried objects within dynamic video segments. Prevailing fully-trained approaches are notorious

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization

DGX agent

arXiv:2604.12887v1 Announce Type: new Abstract: Visual tokenizers map high-dimensional raw pixels into a compressed representation for downstream modeling. Beyond compression, tokenizers dictate what

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching

DGX agent

arXiv:2507.09318v2 Announce Type: replace-cross Abstract: Generating spoken dialogue is inherently more complex than monologue text-to-speech (TTS), as it demands both realistic turn-taking and the ma

model-releasesarxiv-cs-cl
15 Apr 2026
Tutorials

A Data-driven Loss Weighting Scheme across Heterogeneous Tasks for Image Denoising

DGX agent

arXiv:2301.06081v4 Announce Type: replace-cross Abstract: In a variational denoising model, weight in the data fidelity term plays the role of enhancing the noise-removal capability. It is profoundly

tutorialsarxiv-cs-cv
14 Apr 2026
Research

A-IO: Adaptive Inference Orchestration for Memory-Bound NPUs

DGX agent

arXiv:2604.09752v1 Announce Type: cross Abstract: During the deployment of Large Language Models (LLMs), the autoregressive decoding phase on heterogeneous NPU platforms (e.g., Ascend 910B) faces seve

researcharxiv-cs-ai
14 Apr 2026
← Previous
1…383384385386387…1074
Next →