AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,623 results
Research

DPQuant: Efficient and Differentially-Private Model Training via Dynamic Quantization Scheduling

DGX agent

arXiv:2509.03472v2 Announce Type: replace Abstract: Differentially-Private SGD (DP-SGD) and its adaptive variant DP-Adam are powerful techniques to protect user privacy when using sensitive data to tr

researcharxiv-cs-lg
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Language Model as Planner and Formalizer under Constraints

DGX agent

arXiv:2510.05486v2 Announce Type: replace Abstract: LLMs have been widely used in planning, either as planners to generate action sequences end-to-end, or as formalizers to represent the planning doma

safetyarxiv-cs-cl
17 Apr 2026
Applications

OmniLight: One Model to Rule All Lighting Conditions

DGX agent

arXiv:2604.15170v1 Announce Type: new Abstract: Adverse lighting conditions, such as cast shadows and irregular illumination, pose significant challenges to computer vision systems by degrading visibi

applicationsarxiv-cs-cv
17 Apr 2026
Model Releases

SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization

DGX agent

arXiv:2604.07663v2 Announce Type: replace Abstract: The AdamW optimizer, while standard for LLM pretraining, is a critical memory bottleneck, consuming optimizer states equivalent to twice the model's

model-releasesarxiv-cs-lg
17 Apr 2026
Research

Scalable Model-Based Clustering with Sequential Monte Carlo

DGX agent

arXiv:2604.14810v1 Announce Type: cross Abstract: In online clustering problems, there is often a large amount of uncertainty over possible cluster assignments that cannot be resolved until more data

researcharxiv-cs-lg
17 Apr 2026
Research

Unsupervised Learning of Local Updates for Maximum Independent Set in Dynamic Graphs

DGX agent

arXiv:2505.13754v3 Announce Type: replace Abstract: We present the first unsupervised learning model for Maximum-Independent-Set (MaxIS) in dynamic graphs where edges change over time. Our method comb

researcharxiv-cs-lg
17 Apr 2026
Applications

Blind Bitstream-corrupted Video Recovery via Metadata-guided Diffusion Model

DGX agent

arXiv:2604.13906v1 Announce Type: new Abstract: Bitstream-corrupted video recovery aims to restore realistic content degraded during video storage or transmission. Existing methods typically assume th

applicationsarxiv-cs-cv
16 Apr 2026
Research

Caption First, VQA Second: Knowledge Density, Not Task Format, Drives Multimodal Scaling

DGX agent

arXiv:2604.13054v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved rapid progress, yet their scaling behavior remains less clearly characterized and often less pred

researcharxiv-cs-cl
16 Apr 2026
Research

Getting the Numbers Rightnicode{x2014}Modelling Multi-Class Object Counting in Dense and Varied Scenes

DGX agent

arXiv:2510.02213v2 Announce Type: replace Abstract: Density map estimation enables accurate object counting in heavily occluded, and densely packed scenes where detection-based counting fails. In mult

researcharxiv-cs-cv
16 Apr 2026
Research

Graph Propagated Projection Unlearning: A Unified Framework for Vision and Audio Discriminative Models

DGX agent

arXiv:2604.13127v1 Announce Type: new Abstract: The need to selectively and efficiently erase learned information from deep neural networks is becoming increasingly important for privacy, regulatory c

researcharxiv-cs-cv
16 Apr 2026
Model Releases

Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs

DGX agent

arXiv:2604.13258v1 Announce Type: new Abstract: Attribution methods seek to explain language model predictions by quantifying the contribution of input tokens to generated outputs. However, most exist

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Monthly Diffusion v0.9: A Latent Diffusion Model for the First AI-MIP

DGX agent

arXiv:2604.13481v1 Announce Type: new Abstract: Here, we describe Monthly Diffusion at 1.5-degree grid spacing (MD-1.5 version 0.9), a climate emulator that leverages a spherical Fourier neural operat

researcharxiv-cs-lg
16 Apr 2026
Safety

Neural Mean-Field Games: Extending Mean-Field Game Theory with Neural Stochastic Differential Equations

DGX agent

arXiv:2504.13228v4 Announce Type: replace Abstract: Mean-field game theory relies on approximating games that are intractable to model due to a very large to infinite population of players. While thes

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

DGX agent

arXiv:2604.14144v1 Announce Type: cross Abstract: Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

DGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Analyzing the Effect of Noise in LLM Fine-tuning

DGX agent

arXiv:2604.12469v1 Announce Type: new Abstract: Fine-tuning is the dominant paradigm for adapting pretrained large language models (LLMs) to downstream NLP tasks. In practice, fine-tuning datasets may

model-releasesarxiv-cs-lg
15 Apr 2026
Applications

ArtifactWorld: Scaling 3D Gaussian Splatting Artifact Restoration via Video Generation Models

DGX agent

arXiv:2604.12251v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) delivers high-fidelity real-time rendering but suffers from geometric and photometric degradations under sparse-view constr

applicationsarxiv-cs-cv
15 Apr 2026
Safety

ASGuard: Activation-Scaling Guard to Mitigate Targeted Jailbreaking Attack

DGX agent

arXiv:2509.25843v2 Announce Type: replace Abstract: Large language models (LLMs), despite being safety-aligned, exhibit brittle refusal behaviors that can be circumvented by simple linguistic changes.

safetyarxiv-cs-ai
15 Apr 2026
Research

CoLA: A Choice Leakage Attack Framework to Expose Privacy Risks in Subset Training

DGX agent

arXiv:2604.12342v1 Announce Type: cross Abstract: Training models on a carefully chosen portion of data rather than the full dataset is now a standard preprocess for modern ML. From vision coreset sel

researcharxiv-cs-cv
15 Apr 2026
Safety

ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance

DGX agent

arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

Generative Refinement Networks for Visual Synthesis

DGX agent

arXiv:2604.13030v1 Announce Type: new Abstract: While diffusion models dominate the field of visual generation, they are computationally inefficient, applying a uniform computational effort regardless

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Nucleus-Image: Sparse MoE for Image Generation

DGX agent

arXiv:2604.12163v1 Announce Type: new Abstract: We present Nucleus-Image, a text-to-image generation model that establishes a new Pareto frontier in quality-versus-efficiency by matching or exceeding

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

DGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

safetyarxiv-cs-ai
15 Apr 2026
Research

SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis

DGX agent

arXiv:2604.13035v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) increasingly generate indoor scenes through intermediate structures such as layouts and

researcharxiv-cs-cl
15 Apr 2026
Research

SOLARIS: Speculative Offloading of Latent-bAsed Representation for Inference Scaling

DGX agent

arXiv:2604.12110v1 Announce Type: new Abstract: Recent advances in recommendation scaling laws have led to foundation models of unprecedented complexity. While these models offer superior performance,

researcharxiv-cs-lg
15 Apr 2026
Research

Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling

DGX agent

arXiv:2508.04282v3 Announce Type: replace Abstract: Recent benchmarks for memory-augmented reinforcement learning (RL) have introduced partially observable Markov decision process (POMDP) environments

researcharxiv-cs-ai
15 Apr 2026
Model Releases

TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment

DGX agent

arXiv:2604.12012v1 Announce Type: new Abstract: Recent progress in vision-language pretraining has enabled significant improvements to many downstream computer vision applications, such as classificat

model-releasesarxiv-cs-cv
15 Apr 2026
Local Ai

ViLL-E: Video LLM Embeddings for Retrieval

DGX agent

arXiv:2604.12148v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) excel at video understanding tasks where outputs are textual, such as Video Question Answering and Video Caption

local-aiarxiv-cs-cv
15 Apr 2026
Local Ai

A Two-Stage Dual-Modality Model for Facial Emotional Expression Recognition

DGX agent

arXiv:2603.12221v2 Announce Type: replace Abstract: This paper addresses the expression (EXPR) recognition challenge in the 10th Affective Behavior Analysis in-the-Wild (ABAW) workshop and competition

local-aiarxiv-cs-cv
14 Apr 2026
Safety

bioLeak: Leakage-Aware Modeling and Diagnostics for Machine Learning in R

DGX agent

arXiv:2604.10965v1 Announce Type: cross Abstract: Data leakage remains a recurrent source of optimistic bias in biomedical machine learning studies. Standard row-wise cross-validation and globally est

safetyarxiv-cs-lg
14 Apr 2026
Safety

Closed-Form Concept Erasure via Double Projections

DGX agent

arXiv:2604.10032v1 Announce Type: cross Abstract: While modern generative models such as diffusion-based architectures have enabled impressive creative capabilities, they also raise important safety a

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification

DGX agent

arXiv:2604.10054v1 Announce Type: new Abstract: This study addresses a key limitation in deep learning Automatic Modulation Classification (AMC) models, which perform well at high signal-to-noise rati

model-releasesarxiv-cs-lg
14 Apr 2026
Applications

deCIFer: Crystal Structure Prediction from Powder Diffraction Data using Autoregressive Language Models

DGX agent

arXiv:2502.02189v4 Announce Type: replace Abstract: Novel materials drive advancements in fields ranging from energy storage to electronics, with crystal structure characterization forming a crucial y

applicationsarxiv-cs-lg
14 Apr 2026
Agents

Dual-Control Frequency-Aware Diffusion Model for Depth-Dependent Optical Microrobot Microscopy Image Generation

DGX agent

arXiv:2604.11680v1 Announce Type: new Abstract: Optical microrobots actuated by optical tweezers (OT) are important for cell manipulation and microscale assembly, but their autonomous operation depend

agentsarxiv-cs-ro
14 Apr 2026
Safety

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models

DGX agent

arXiv:2604.11043v1 Announce Type: new Abstract: Unified multimodal embedding spaces underpin practical applications such as cross-modal retrieval and zero-shot recognition. In many real deployments, h

safetyarxiv-cs-ai
14 Apr 2026
Safety

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering

DGX agent

arXiv:2511.09282v3 Announce Type: replace-cross Abstract: Significant progress has been made in spoken question answering (SQA) in recent years. However, many existing methods, including large audio l

safetyarxiv-cs-cl
14 Apr 2026
Research

H-SPAM: Hierarchical Superpixel Anything Model

DGX agent

arXiv:2604.11218v1 Announce Type: new Abstract: Superpixels offer a compact image representation by grouping pixels into coherent regions. Recent methods have reached a plateau in terms of segmentatio

researcharxiv-cs-cv
14 Apr 2026
Model Releases

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

DGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

DGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

model-releasesnous-research--x
14 Apr 2026
Hardware

IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7

DGX agent

This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The

hardwarer-ollama
14 Apr 2026
Model Releases

Infusing Theory of Mind into Socially Intelligent LLM Agents

DGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

MASH: Modeling Abstention via Selective Help-Seeking

DGX agent

arXiv:2510.01152v2 Announce Type: replace Abstract: LLMs cannot reliably recognize their parametric knowledge boundaries and often hallucinate answers to outside-of-boundary questions. In this paper,

agentsarxiv-cs-cl
14 Apr 2026
Safety

MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration

DGX agent

arXiv:2604.10081v1 Announce Type: cross Abstract: Real-world image pairs often exhibit both severe degradations and large viewpoint changes, making image restoration and geometric matching mutually in

safetyarxiv-cs-ai
14 Apr 2026
Research

MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models

DGX agent

arXiv:2408.11871v4 Announce Type: replace-cross Abstract: Fake news significantly influences decision-making processes by misleading individuals, organizations, and even governments. Large language mo

researcharxiv-cs-ai
14 Apr 2026
Model Releases

MemDLM: Memory-Enhanced DLM Training

DGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo

DGX agent

arXiv:2505.09368v2 Announce Type: replace Abstract: Standard benchmarks for optical flow, scene flow, and stereo vision algorithms generally focus on model accuracy rather than robustness to image cor

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

DGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Thought Branches: Interpreting LLM Reasoning Requires Resampling

DGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

safetyarxiv-cs-ai
14 Apr 2026
← Previous
1…315316317318319…1326
Next →