AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Model Releases

Fundus Image-based Glaucoma Screening via Retinal Knowledge-Oriented Dynamic Multi-Level Feature Integration

DGX agent

arXiv:2604.12351v1 Announce Type: new Abstract: Automated diagnosis based on color fundus photography is essential for large-scale glaucoma screening. However, existing deep learning models are typica

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Gemini 3.1 Flash TTS

DGX agent

Gemini 3.1 Flash TTS Google released Gemini 3.1 Flash TTS today, a new text-to-speech model that can be directed using prompts. It's presented via the standard Gemini API using gemini-3.1-flash-tts-pr

model-releasessimon-willison
15 Apr 2026
Model Releases

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio. Whether you’…

DGX agent

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio. Whether you’re creating a pitch deck or recording a passion project, tra

model-releasesgoogle-ai--x
15 Apr 2026
Safety

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

DGX agent

arXiv:2604.12630v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Rec

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

GTCN-G: A Residual Graph-Temporal Fusion Network for Imbalanced Intrusion Detection

DGX agent

arXiv:2510.07285v3 Announce Type: replace-cross Abstract: The escalating complexity of network threats and the inherent class imbalance in traffic data present formidable challenges for modern Intrusi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Have found the same things! Using glm-5 as a daily driver for a lot of things

DGX agent

Have found the same things! Using glm-5 as a daily driver for a lot of things We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than

model-releasesharrison-chase--x
15 Apr 2026
Model Releases

IDEA: An Interpretable and Editable Decision-Making Framework for LLMs via Verbal-to-Numeric Calibration

DGX agent

arXiv:2604.12573v1 Announce Type: new Abstract: Large Language Models are increasingly deployed for decision-making, yet their adoption in high-stakes domains remains limited by miscalibrated probabil

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study

DGX agent

arXiv:2604.12337v1 Announce Type: new Abstract: Letters of recommendation (LoRs) can carry patterns of implicitly gendered language that can inadvertently influence downstream decisions, e.g. in hirin

model-releasesarxiv-cs-lg
15 Apr 2026
Research

IMU: Influence-guided Machine Unlearning

DGX agent

arXiv:2508.01620v3 Announce Type: replace-cross Abstract: Machine Unlearning (MU) aims to selectively erase the influence of specific data points from pretrained models. However, most existing MU meth

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Invariant Features for Global Crop Type Classification

DGX agent

arXiv:2509.03497v3 Announce Type: replace Abstract: Accurate global crop type mapping supports agricultural monitoring and food security, yet remains limited by the scarcity of labeled data in many re

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Inverse Design of Inorganic Compounds with Generative AI

DGX agent

arXiv:2604.11827v1 Announce Type: cross Abstract: Machine learning is revolutionizing chemistry. Beyond the value of predictive models accelerating virtual screening, generative AI aims at enabling in

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidance

DGX agent

arXiv:2604.12627v1 Announce Type: new Abstract: RLVR improves reasoning in large language models, but its effectiveness is often limited by severe reward sparsity on hard problems. Recent hint-based R

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Leveraging Weighted Syntactic and Semantic Context Assessment Summary (wSSAS) Towards Text Categorization Using LLMs

DGX agent

arXiv:2604.12049v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) for reliable, enterprise-grade analytics such as text categorization is often hindered by the stochastic natur

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

LLM-Guided Prompt Evolution for Password Guessing

DGX agent

arXiv:2604.12601v1 Announce Type: cross Abstract: Passwords still remain a dominant authentication method, yet their security is routinely subverted by predictable user choices and large-scale credent

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

MedConcept: Unsupervised Concept Discovery for Interpretability in Medical VLMs

DGX agent

arXiv:2604.11868v1 Announce Type: new Abstract: While medical Vision-Language models (VLMs) achieve strong performance on tasks such as tumor or organ segmentation and diagnosis prediction, their opaq

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Mema: Memory-Augmented Adapter for Enhanced Vision-Language Understanding

DGX agent

arXiv:2603.00655v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance by aligning pretrained visual representations with the linguistic know

model-releasesarxiv-cs-cv
15 Apr 2026
Research

Mid-training is essential for LLM reasoning, IBM study shows

DGX agent

IBM Research has demonstrated that mid-training — a dedicated training phase between initial pre-training and fine-tuning — is essential for improving reasoning capabilities in large language models (

researchibm-research
15 Apr 2026
Local Ai

More details: https://blog.comfy.org/p/ernie-image-day-0-support

DGX agent

ComfyUI announced day-zero support for ERNIE Image, Baidu's image generation model, with full implementation details available on their official blog. The integration allows users to run ERNIE Image m

local-aicomfyui--x
15 Apr 2026
Local Ai

NucleusAI/Nucleus-Image (Samples)

DGX agent

The Reddit post about NucleusAI/Nucleus-Image appears to be a community showcase of sample outputs from a text-to-image generative AI model called Nucleus-Image, developed by NucleusAI. The post is sh

local-air-stablediffusion
15 Apr 2026
Model Releases

Orthogonal Subspace Projection for Continual Machine Unlearning via SVD-Based LoRA

DGX agent

arXiv:2604.12526v1 Announce Type: cross Abstract: Continual machine unlearning aims to remove the influence of data that should no longer be retained, while preserving the usefulness of the model on e

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Parametric Interpolation of Dynamic Mode Decomposition for Predicting Nonlinear Systems

DGX agent

arXiv:2604.12103v1 Announce Type: cross Abstract: We present parameter-interpolated dynamic mode decomposition (piDMD), a parametric reduced-order modeling framework that embeds known parameter-affine

model-releasesarxiv-cs-lg
15 Apr 2026
Hardware

PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving

DGX agent

arXiv:2604.12171v1 Announce Type: cross Abstract: Pipeline parallelism (PP) is widely used to partition layers of large language models (LLMs) across GPUs, enabling scalable inference for large models

hardwarearxiv-cs-lg
15 Apr 2026
Model Releases

ReflectCAP: Detailed Image Captioning with Reflective Memory

DGX agent

arXiv:2604.12357v1 Announce Type: new Abstract: Detailed image captioning demands both factual grounding and fine-grained coverage, yet existing methods have struggled to achieve them simultaneously.

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Should There be a Teacher In-the-Loop? A Study of Generative AI Personalized Tasks Middle School

DGX agent

arXiv:2602.15876v1 Announce Type: cross Abstract: Adapting instruction to the fine-grained needs of individual students is a powerful application of recent advances in large language models. These gen

researcharxiv-cs-ai
15 Apr 2026
Model Releases

SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks

DGX agent

arXiv:2506.14512v4 Announce Type: replace Abstract: Large Language Models (LLMs) have undergone rapid progress, largely attributed to reinforcement learning on complex reasoning tasks. In contrast, wh

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

Spatial-Spectral Adaptive Fidelity and Noise Prior Reduction Guided Hyperspectral Image Denoising

DGX agent

arXiv:2604.12600v1 Announce Type: new Abstract: The core challenge of hyperspectral image denoising is striking the right balance between data fidelity and noise prior modeling. Most existing methods

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

The Verification Tax: Fundamental Limits of AI Auditing in the Rare-Error Regime

DGX agent

arXiv:2604.12951v1 Announce Type: new Abstract: The most cited calibration result in deep learning -- post-temperature-scaling ECE of 0.012 on CIFAR-100 (Guo et al., 2017) -- is below the statistical

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Thought-Retriever: Don't Just Retrieve Raw Data, Retrieve Thoughts for Memory-Augmented Agentic Systems

DGX agent

arXiv:2604.12231v1 Announce Type: new Abstract: Large language models (LLMs) have transformed AI research thanks to their powerful internal capabilities and knowledge. However, existing LLMs still fai

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Unlocking the Potential of Grounding DINO in Videos: Parameter-Efficient Adaptation for Limited-Data Spatial-Temporal Localization

DGX agent

arXiv:2604.12346v1 Announce Type: new Abstract: Spatio-temporal video grounding (STVG) aims to localize queried objects within dynamic video segments. Prevailing fully-trained approaches are notorious

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization

DGX agent

arXiv:2604.12887v1 Announce Type: new Abstract: Visual tokenizers map high-dimensional raw pixels into a compressed representation for downstream modeling. Beyond compression, tokenizers dictate what

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching

DGX agent

arXiv:2507.09318v2 Announce Type: replace-cross Abstract: Generating spoken dialogue is inherently more complex than monologue text-to-speech (TTS), as it demands both realistic turn-taking and the ma

model-releasesarxiv-cs-cl
15 Apr 2026
Tutorials

A Data-driven Loss Weighting Scheme across Heterogeneous Tasks for Image Denoising

DGX agent

arXiv:2301.06081v4 Announce Type: replace-cross Abstract: In a variational denoising model, weight in the data fidelity term plays the role of enhancing the noise-removal capability. It is profoundly

tutorialsarxiv-cs-cv
14 Apr 2026
Research

A-IO: Adaptive Inference Orchestration for Memory-Bound NPUs

DGX agent

arXiv:2604.09752v1 Announce Type: cross Abstract: During the deployment of Large Language Models (LLMs), the autoregressive decoding phase on heterogeneous NPU platforms (e.g., Ascend 910B) faces seve

researcharxiv-cs-ai
14 Apr 2026
Model Releases

AI-enhanced tuning of quantum dot Hamiltonians toward Majorana modes

DGX agent

arXiv:2601.02149v3 Announce Type: replace-cross Abstract: We propose a neural network-based model capable of learning the broad landscape of working regimes in quantum dot simulators, and using this k

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

AIM-Bench: Benchmarking and Improving Affective Image Manipulation via Fine-Grained Hierarchical Control

DGX agent

arXiv:2604.10454v1 Announce Type: new Abstract: Affective Image Manipulation (AIM) aims to evoke specific emotions through targeted editing. Current image editing benchmarks primarily focus on object-

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

AIMER: Calibration-Free Task-Agnostic MoE Pruning

DGX agent

arXiv:2603.18492v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) language models increase parameter capacity without proportional per-token compute, but the deployment still requires stori

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs

DGX agent

arXiv:2604.10528v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) demonstrate remarkable zero-shot recognition capabilities across a diverse spectrum of multimodal tasks, it yet rema

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs

DGX agent

arXiv:2604.10585v1 Announce Type: cross Abstract: Modern large language models (LLMs) are increasingly fine-tuned via reinforcement learning from human feedback (RLHF) or related reward optimisation s

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

DGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

model-releasesarxiv-cs-lg
14 Apr 2026
Local Ai

ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection

DGX agent

arXiv:2604.11790v1 Announce Type: cross Abstract: Tool-augmented Large Language Model (LLM) agents have demonstrated impressive capabilities in automating complex, multi-step real-world tasks, yet rem

local-aiarxiv-cs-ai
14 Apr 2026
Research

CodaRAG: Connecting the Dots with Associativity Inspired by Complementary Learning

DGX agent

arXiv:2604.10426v1 Announce Type: cross Abstract: Large Language Models (LLMs) struggle with knowledge-intensive tasks due to hallucinations and fragmented reasoning over dispersed information. While

researcharxiv-cs-ai
14 Apr 2026
Hardware

CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts

DGX agent

arXiv:2604.10496v1 Announce Type: new Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) ar

hardwarearxiv-cs-lg
14 Apr 2026
Local Ai

Correct me if I’m wrong: Ollama can’t fine tune like Unsloth Studio

DGX agent

Ollama is a local inference engine designed for running pre-built LLMs on your own machine, and it does not include fine-tuning capabilities — this distinction is correct. Unsloth (and its Unsloth Stu

local-air-ollama
14 Apr 2026
Model Releases

CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing

DGX agent

arXiv:2506.18438v2 Announce Type: replace Abstract: Editing natural images using textual descriptions in text-to-image diffusion models remains a significant challenge, particularly in achieving consi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics

DGX agent

arXiv:2512.21877v3 Announce Type: replace-cross Abstract: Cricket is the second most popular sport worldwide, with billions of fans seeking advanced statistical insights unavailable through standard w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

DecepGPT: Schema-Driven Deception Detection with Multicultural Datasets and Robust Multimodal Learning

DGX agent

arXiv:2603.23916v2 Announce Type: replace-cross Abstract: Multimodal deception detection aims to identify deceptive behavior by analyzing audiovisual cues for forensics and security. In these high-sta

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Defending against Backdoor Attacks via Module Switching

DGX agent

arXiv:2504.05902v2 Announce Type: replace-cross Abstract: Backdoor attacks pose a serious threat to deep neural networks (DNNs), allowing adversaries to implant triggers for hidden behaviors in infere

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Delta Rectified Flow Sampling for Text-to-Image Editing

DGX agent

arXiv:2509.05342v3 Announce Type: replace Abstract: We propose Delta Rectified Flow Sampling (DRFS), a novel inversion-free, path-aware editing framework within rectified flow models for text-to-image

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…478479480481482…1358
Next →