AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlog
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,458 results
Safety

ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance

DGX agent

arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires

safetyarxiv-cs-cl
15 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Generative Refinement Networks for Visual Synthesis

DGX agent

arXiv:2604.13030v1 Announce Type: new Abstract: While diffusion models dominate the field of visual generation, they are computationally inefficient, applying a uniform computational effort regardless

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Nucleus-Image: Sparse MoE for Image Generation

DGX agent

arXiv:2604.12163v1 Announce Type: new Abstract: We present Nucleus-Image, a text-to-image generation model that establishes a new Pareto frontier in quality-versus-efficiency by matching or exceeding

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

DGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

safetyarxiv-cs-ai
15 Apr 2026
Research

SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis

DGX agent

arXiv:2604.13035v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) increasingly generate indoor scenes through intermediate structures such as layouts and

researcharxiv-cs-cl
15 Apr 2026
Research

SOLARIS: Speculative Offloading of Latent-bAsed Representation for Inference Scaling

DGX agent

arXiv:2604.12110v1 Announce Type: new Abstract: Recent advances in recommendation scaling laws have led to foundation models of unprecedented complexity. While these models offer superior performance,

researcharxiv-cs-lg
15 Apr 2026
Research

Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling

DGX agent

arXiv:2508.04282v3 Announce Type: replace Abstract: Recent benchmarks for memory-augmented reinforcement learning (RL) have introduced partially observable Markov decision process (POMDP) environments

researcharxiv-cs-ai
15 Apr 2026
Model Releases

TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment

DGX agent

arXiv:2604.12012v1 Announce Type: new Abstract: Recent progress in vision-language pretraining has enabled significant improvements to many downstream computer vision applications, such as classificat

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

ViLL-E: Video LLM Embeddings for Retrieval

DGX agent

arXiv:2604.12148v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) excel at video understanding tasks where outputs are textual, such as Video Question Answering and Video Caption

local-aiarxiv-cs-cv
15 Apr 2026
Local Ai

A Two-Stage Dual-Modality Model for Facial Emotional Expression Recognition

DGX agent

arXiv:2603.12221v2 Announce Type: replace Abstract: This paper addresses the expression (EXPR) recognition challenge in the 10th Affective Behavior Analysis in-the-Wild (ABAW) workshop and competition

local-aiarxiv-cs-cv
14 Apr 2026
Safety

bioLeak: Leakage-Aware Modeling and Diagnostics for Machine Learning in R

DGX agent

arXiv:2604.10965v1 Announce Type: cross Abstract: Data leakage remains a recurrent source of optimistic bias in biomedical machine learning studies. Standard row-wise cross-validation and globally est

safetyarxiv-cs-lg
14 Apr 2026
Safety

Closed-Form Concept Erasure via Double Projections

DGX agent

arXiv:2604.10032v1 Announce Type: cross Abstract: While modern generative models such as diffusion-based architectures have enabled impressive creative capabilities, they also raise important safety a

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification

DGX agent

arXiv:2604.10054v1 Announce Type: new Abstract: This study addresses a key limitation in deep learning Automatic Modulation Classification (AMC) models, which perform well at high signal-to-noise rati

model-releasesarxiv-cs-lg
14 Apr 2026
Applications

deCIFer: Crystal Structure Prediction from Powder Diffraction Data using Autoregressive Language Models

DGX agent

arXiv:2502.02189v4 Announce Type: replace Abstract: Novel materials drive advancements in fields ranging from energy storage to electronics, with crystal structure characterization forming a crucial y

applicationsarxiv-cs-lg
14 Apr 2026
Agents

Dual-Control Frequency-Aware Diffusion Model for Depth-Dependent Optical Microrobot Microscopy Image Generation

DGX agent

arXiv:2604.11680v1 Announce Type: new Abstract: Optical microrobots actuated by optical tweezers (OT) are important for cell manipulation and microscale assembly, but their autonomous operation depend

agentsarxiv-cs-ro
14 Apr 2026
Safety

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models

DGX agent

arXiv:2604.11043v1 Announce Type: new Abstract: Unified multimodal embedding spaces underpin practical applications such as cross-modal retrieval and zero-shot recognition. In many real deployments, h

safetyarxiv-cs-ai
14 Apr 2026
Safety

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering

DGX agent

arXiv:2511.09282v3 Announce Type: replace-cross Abstract: Significant progress has been made in spoken question answering (SQA) in recent years. However, many existing methods, including large audio l

safetyarxiv-cs-cl
14 Apr 2026
Research

H-SPAM: Hierarchical Superpixel Anything Model

DGX agent

arXiv:2604.11218v1 Announce Type: new Abstract: Superpixels offer a compact image representation by grouping pixels into coherent regions. Recent methods have reached a plateau in terms of segmentatio

researcharxiv-cs-cv
14 Apr 2026
Model Releases

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

DGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

DGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

model-releasesnous-research--x
14 Apr 2026
Hardware

IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7

DGX agent

This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The

hardwarer-ollama
14 Apr 2026
Model Releases

Infusing Theory of Mind into Socially Intelligent LLM Agents

DGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

MASH: Modeling Abstention via Selective Help-Seeking

DGX agent

arXiv:2510.01152v2 Announce Type: replace Abstract: LLMs cannot reliably recognize their parametric knowledge boundaries and often hallucinate answers to outside-of-boundary questions. In this paper,

agentsarxiv-cs-cl
14 Apr 2026
Safety

MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration

DGX agent

arXiv:2604.10081v1 Announce Type: cross Abstract: Real-world image pairs often exhibit both severe degradations and large viewpoint changes, making image restoration and geometric matching mutually in

safetyarxiv-cs-ai
14 Apr 2026
Research

MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models

DGX agent

arXiv:2408.11871v4 Announce Type: replace-cross Abstract: Fake news significantly influences decision-making processes by misleading individuals, organizations, and even governments. Large language mo

researcharxiv-cs-ai
14 Apr 2026
Model Releases

MemDLM: Memory-Enhanced DLM Training

DGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo

DGX agent

arXiv:2505.09368v2 Announce Type: replace Abstract: Standard benchmarks for optical flow, scene flow, and stereo vision algorithms generally focus on model accuracy rather than robustness to image cor

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

DGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Thought Branches: Interpreting LLM Reasoning Requires Resampling

DGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

safetyarxiv-cs-ai
14 Apr 2026
Agents

Towards Autonomous Mechanistic Reasoning in Virtual Cells

DGX agent

arXiv:2604.11661v1 Announce Type: cross Abstract: Large language models (LLMs) have recently gained significant attention as a promising approach to accelerate scientific discovery. However, their app

agentsarxiv-cs-ai
14 Apr 2026
Research

Zero-shot World Models Are Developmentally Efficient Learners

DGX agent

arXiv:2604.10333v1 Announce Type: new Abstract: Young children demonstrate early abilities to understand their physical world, estimating depth, motion, object coherence, interactions, and many other

researcharxiv-cs-ai
14 Apr 2026
Local Ai

Another BRIXEL in the Wall: Towards Cheaper Dense Features

DGX agent

arXiv:2511.05168v2 Announce Type: replace Abstract: Vision foundation models achieve strong performance on both global and locally dense downstream tasks. Pretrained on large images, the recent DINOv3

local-aiarxiv-cs-cv
13 Apr 2026
Research

Beyond Isolated Clients: Integrating Graph-Based Embeddings into Event Sequence Models

DGX agent

arXiv:2604.09085v1 Announce Type: cross Abstract: Large-scale digital platforms generate billions of timestamped user-item interactions (events) that are crucial for predicting user attributes in, e.g

researcharxiv-cs-ai
13 Apr 2026
Applications

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models

DGX agent

arXiv:2604.06161v2 Announce Type: replace-cross Abstract: Most digital videos are stored in 8-bit low dynamic range (LDR) formats, where much of the original high dynamic range (HDR) scene radiance is

applicationsarxiv-cs-ai
13 Apr 2026
Research

Efficient Unlearning through Maximizing Relearning Convergence Delay

DGX agent

arXiv:2604.09391v1 Announce Type: cross Abstract: Machine unlearning poses challenges in removing mislabeled, contaminated, or problematic data from a pretrained model. Current unlearning approaches a

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Gemma:26b thinking issue in openWebUI

DGX agent

This r/ollama thread discusses user-reported issues with the Gemma 4 26B (a Mixture of Experts model) and its 'thinking' mode when used through Open WebUI. Key problems include the model getting stuck

model-releasesr-ollama
13 Apr 2026
Safety

InstrAct: Towards Action-Centric Understanding in Instructional Videos

DGX agent

arXiv:2604.08762v1 Announce Type: cross Abstract: Understanding instructional videos requires recognizing fine-grained actions and modeling their temporal relations, which remains challenging for curr

safetyarxiv-cs-ai
13 Apr 2026
Research

Interactive Program Synthesis for Modeling Collaborative Physical Activities from Narrated Demonstrations

DGX agent

arXiv:2509.24250v3 Announce Type: replace Abstract: Teaching systems physical tasks is a long standing goal in HCI, yet most prior work has focused on non collaborative physical activities. Collaborat

researcharxiv-cs-ai
13 Apr 2026
Research

Large-Scale Universal Defect Generation: Foundation Models and Datasets

DGX agent

arXiv:2604.08915v1 Announce Type: cross Abstract: Existing defect/anomaly generation methods often rely on few-shot learning, which overfits to specific defect categories due to the lack of large-scal

researcharxiv-cs-ai
13 Apr 2026
Research

LLM4Delay: Flight Delay Prediction via Cross-Modality Adaptation of Large Language Models and Aircraft Trajectory Representation

DGX agent

arXiv:2510.23636v3 Announce Type: replace-cross Abstract: Flight delay prediction has become a key focus in air traffic management (ATM), as delays reflect inefficiencies in the system. This paper pro

researcharxiv-cs-ai
13 Apr 2026
Local Ai

Need help to download from civitai in China

DGX agent

This Reddit thread from r/StableDiffusion addresses the challenge faced by users in China trying to access and download models from Civitai, which may be restricted or slow due to network limitations

local-air-stablediffusion
13 Apr 2026
Model Releases

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

DGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

model-releasesarxiv-cs-ai
13 Apr 2026
Tutorials

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

DGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

tutorialsarxiv-cs-ai
13 Apr 2026
Hardware

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

DGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

hardwarearxiv-cs-ai
10 Apr 2026
Model Releases

A-MBER: Affective Memory Benchmark for Emotion Recognition

DGX agent

arXiv:2604.07017v1 Announce Type: new Abstract: AI assistants that interact with users over time need to interpret the user's current emotional state in order to respond appropriately and personally.

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Ace Step 1.5 XL ComfyUI automation workflow without lama for generating random tags using qwen, generate song and then give it a rating by using waveform analysis

DGX agent

This is a ComfyUI automation workflow for the ACE-Step 1.5 XL music generation model that uses Qwen (a language model text encoder) to randomly generate music tags/captions without requiring LAMA, ...

model-releasesr-stablediffusion
10 Apr 2026
Research

Adapting Foundation Models for Annotation-Efficient Adnexal Mass Segmentation in Cine Images

DGX agent

arXiv:2604.08045v1 Announce Type: new Abstract: Adnexal mass evaluation via ultrasound is a challenging clinical task, often hindered by subjective interpretation and significant inter-observer variab

researcharxiv-cs-cv
10 Apr 2026
Model Releases

AgriPath: A Systematic Exploration of Architectural Trade-offs for Crop Disease Classification

DGX agent

arXiv:2603.13354v3 Announce Type: replace-cross Abstract: Reliable crop disease detection requires models that perform consistently across diverse acquisition conditions, yet existing evaluations ofte

model-releasesarxiv-cs-lg
10 Apr 2026
← Previous
1…309310311312313…1302
Next →