AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
17 Apr 2026

BitFlipScope: Scalable Fault Localization and Recovery for Bit-Flip Corruptions in LLMs

Local AiDGX agent

arXiv:2512.22174v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) deployed in practical and safety-critical settings are increasingly susceptible to bit-flip faults caused by hard

Bridging the Gap between Learning and Inference for Diffusion-Based Molecule Generation

Model ReleasesDGX agent

arXiv:2411.05472v2 Announce Type: replace Abstract: The paradigm shift toward structure-driven molecule generation has been propelled by advances in deep generative models, such as variational auto-en

Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades

SafetyDGX agent

arXiv:2604.14251v1 Announce Type: new Abstract: Monitoring LLM safety at scale requires balancing cost and accuracy: a cheap latent-space probe can screen every input, but hard cases should be escalat

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Can LLMs Score Medical Diagnoses and Clinical Reasoning as well as Expert Panels?

SafetyDGX agent

arXiv:2604.14892v1 Announce Type: new Abstract: Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudi

DeepPrune: Parallel Scaling without Inter-trace Redundancy

ResearchDGX agent

arXiv:2510.08483v2 Announce Type: replace Abstract: Parallel scaling has emerged as a powerful paradigm to enhance reasoning capabilities in large language models (LLMs) by generating multiple Chain-o

DPQuant: Efficient and Differentially-Private Model Training via Dynamic Quantization Scheduling

ResearchDGX agent

arXiv:2509.03472v2 Announce Type: replace Abstract: Differentially-Private SGD (DP-SGD) and its adaptive variant DP-Adam are powerful techniques to protect user privacy when using sensitive data to tr

Language Model as Planner and Formalizer under Constraints

SafetyDGX agent

arXiv:2510.05486v2 Announce Type: replace Abstract: LLMs have been widely used in planning, either as planners to generate action sequences end-to-end, or as formalizers to represent the planning doma

OmniLight: One Model to Rule All Lighting Conditions

ApplicationsDGX agent

arXiv:2604.15170v1 Announce Type: new Abstract: Adverse lighting conditions, such as cast shadows and irregular illumination, pose significant challenges to computer vision systems by degrading visibi

SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization

Model ReleasesDGX agent

arXiv:2604.07663v2 Announce Type: replace Abstract: The AdamW optimizer, while standard for LLM pretraining, is a critical memory bottleneck, consuming optimizer states equivalent to twice the model's

Scalable Model-Based Clustering with Sequential Monte Carlo

ResearchDGX agent

arXiv:2604.14810v1 Announce Type: cross Abstract: In online clustering problems, there is often a large amount of uncertainty over possible cluster assignments that cannot be resolved until more data

Unsupervised Learning of Local Updates for Maximum Independent Set in Dynamic Graphs

ResearchDGX agent

arXiv:2505.13754v3 Announce Type: replace Abstract: We present the first unsupervised learning model for Maximum-Independent-Set (MaxIS) in dynamic graphs where edges change over time. Our method comb

16 Apr 2026

Blind Bitstream-corrupted Video Recovery via Metadata-guided Diffusion Model

ApplicationsDGX agent

arXiv:2604.13906v1 Announce Type: new Abstract: Bitstream-corrupted video recovery aims to restore realistic content degraded during video storage or transmission. Existing methods typically assume th

Caption First, VQA Second: Knowledge Density, Not Task Format, Drives Multimodal Scaling

ResearchDGX agent

arXiv:2604.13054v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved rapid progress, yet their scaling behavior remains less clearly characterized and often less pred

Getting the Numbers Rightnicode{x2014}Modelling Multi-Class Object Counting in Dense and Varied Scenes

ResearchDGX agent

arXiv:2510.02213v2 Announce Type: replace Abstract: Density map estimation enables accurate object counting in heavily occluded, and densely packed scenes where detection-based counting fails. In mult

Graph Propagated Projection Unlearning: A Unified Framework for Vision and Audio Discriminative Models

ResearchDGX agent

arXiv:2604.13127v1 Announce Type: new Abstract: The need to selectively and efficiently erase learned information from deep neural networks is becoming increasingly important for privacy, regulatory c

Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs

Model ReleasesDGX agent

arXiv:2604.13258v1 Announce Type: new Abstract: Attribution methods seek to explain language model predictions by quantifying the contribution of input tokens to generated outputs. However, most exist

Monthly Diffusion v0.9: A Latent Diffusion Model for the First AI-MIP

ResearchDGX agent

arXiv:2604.13481v1 Announce Type: new Abstract: Here, we describe Monthly Diffusion at 1.5-degree grid spacing (MD-1.5 version 0.9), a climate emulator that leverages a spherical Fourier neural operat

Neural Mean-Field Games: Extending Mean-Field Game Theory with Neural Stochastic Differential Equations

SafetyDGX agent

arXiv:2504.13228v4 Announce Type: replace Abstract: Mean-field game theory relies on approximating games that are intractable to model due to a very large to infinite population of players. While thes

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

Model ReleasesDGX agent

arXiv:2604.14144v1 Announce Type: cross Abstract: Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

Model ReleasesDGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

15 Apr 2026

Analyzing the Effect of Noise in LLM Fine-tuning

Model ReleasesDGX agent

arXiv:2604.12469v1 Announce Type: new Abstract: Fine-tuning is the dominant paradigm for adapting pretrained large language models (LLMs) to downstream NLP tasks. In practice, fine-tuning datasets may

ArtifactWorld: Scaling 3D Gaussian Splatting Artifact Restoration via Video Generation Models

ApplicationsDGX agent

arXiv:2604.12251v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) delivers high-fidelity real-time rendering but suffers from geometric and photometric degradations under sparse-view constr

ASGuard: Activation-Scaling Guard to Mitigate Targeted Jailbreaking Attack

SafetyDGX agent

arXiv:2509.25843v2 Announce Type: replace Abstract: Large language models (LLMs), despite being safety-aligned, exhibit brittle refusal behaviors that can be circumvented by simple linguistic changes.

CoLA: A Choice Leakage Attack Framework to Expose Privacy Risks in Subset Training

ResearchDGX agent

arXiv:2604.12342v1 Announce Type: cross Abstract: Training models on a carefully chosen portion of data rather than the full dataset is now a standard preprocess for modern ML. From vision coreset sel

ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance

SafetyDGX agent

arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires

Generative Refinement Networks for Visual Synthesis

Model ReleasesDGX agent

arXiv:2604.13030v1 Announce Type: new Abstract: While diffusion models dominate the field of visual generation, they are computationally inefficient, applying a uniform computational effort regardless

Nucleus-Image: Sparse MoE for Image Generation

Model ReleasesDGX agent

arXiv:2604.12163v1 Announce Type: new Abstract: We present Nucleus-Image, a text-to-image generation model that establishes a new Pareto frontier in quality-versus-efficiency by matching or exceeding

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

SafetyDGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis

ResearchDGX agent

arXiv:2604.13035v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) increasingly generate indoor scenes through intermediate structures such as layouts and

SOLARIS: Speculative Offloading of Latent-bAsed Representation for Inference Scaling

ResearchDGX agent

arXiv:2604.12110v1 Announce Type: new Abstract: Recent advances in recommendation scaling laws have led to foundation models of unprecedented complexity. While these models offer superior performance,

Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling

ResearchDGX agent

arXiv:2508.04282v3 Announce Type: replace Abstract: Recent benchmarks for memory-augmented reinforcement learning (RL) have introduced partially observable Markov decision process (POMDP) environments

TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment

Model ReleasesDGX agent

arXiv:2604.12012v1 Announce Type: new Abstract: Recent progress in vision-language pretraining has enabled significant improvements to many downstream computer vision applications, such as classificat

ViLL-E: Video LLM Embeddings for Retrieval

Local AiDGX agent

arXiv:2604.12148v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) excel at video understanding tasks where outputs are textual, such as Video Question Answering and Video Caption

14 Apr 2026

A Two-Stage Dual-Modality Model for Facial Emotional Expression Recognition

Local AiDGX agent

arXiv:2603.12221v2 Announce Type: replace Abstract: This paper addresses the expression (EXPR) recognition challenge in the 10th Affective Behavior Analysis in-the-Wild (ABAW) workshop and competition

bioLeak: Leakage-Aware Modeling and Diagnostics for Machine Learning in R

SafetyDGX agent

arXiv:2604.10965v1 Announce Type: cross Abstract: Data leakage remains a recurrent source of optimistic bias in biomedical machine learning studies. Standard row-wise cross-validation and globally est

Closed-Form Concept Erasure via Double Projections

SafetyDGX agent

arXiv:2604.10032v1 Announce Type: cross Abstract: While modern generative models such as diffusion-based architectures have enabled impressive creative capabilities, they also raise important safety a

Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification

Model ReleasesDGX agent

arXiv:2604.10054v1 Announce Type: new Abstract: This study addresses a key limitation in deep learning Automatic Modulation Classification (AMC) models, which perform well at high signal-to-noise rati

deCIFer: Crystal Structure Prediction from Powder Diffraction Data using Autoregressive Language Models

ApplicationsDGX agent

arXiv:2502.02189v4 Announce Type: replace Abstract: Novel materials drive advancements in fields ranging from energy storage to electronics, with crystal structure characterization forming a crucial y

Dual-Control Frequency-Aware Diffusion Model for Depth-Dependent Optical Microrobot Microscopy Image Generation

AgentsDGX agent

arXiv:2604.11680v1 Announce Type: new Abstract: Optical microrobots actuated by optical tweezers (OT) are important for cell manipulation and microscale assembly, but their autonomous operation depend

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models

SafetyDGX agent

arXiv:2604.11043v1 Announce Type: new Abstract: Unified multimodal embedding spaces underpin practical applications such as cross-modal retrieval and zero-shot recognition. In many real deployments, h

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering

SafetyDGX agent

arXiv:2511.09282v3 Announce Type: replace-cross Abstract: Significant progress has been made in spoken question answering (SQA) in recent years. However, many existing methods, including large audio l

H-SPAM: Hierarchical Superpixel Anything Model

ResearchDGX agent

arXiv:2604.11218v1 Announce Type: new Abstract: Superpixels offer a compact image representation by grouping pixels into coherent regions. Recent methods have reached a plateau in terms of segmentatio

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

Model ReleasesDGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

Model ReleasesDGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7

HardwareDGX agent

This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The

Infusing Theory of Mind into Socially Intelligent LLM Agents

Model ReleasesDGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

MASH: Modeling Abstention via Selective Help-Seeking

AgentsDGX agent

arXiv:2510.01152v2 Announce Type: replace Abstract: LLMs cannot reliably recognize their parametric knowledge boundaries and often hallucinate answers to outside-of-boundary questions. In this paper,

MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration

SafetyDGX agent

arXiv:2604.10081v1 Announce Type: cross Abstract: Real-world image pairs often exhibit both severe degradations and large viewpoint changes, making image restoration and geometric matching mutually in

MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models

ResearchDGX agent

arXiv:2408.11871v4 Announce Type: replace-cross Abstract: Fake news significantly influences decision-making processes by misleading individuals, organizations, and even governments. Large language mo

MemDLM: Memory-Enhanced DLM Training

Model ReleasesDGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo

Model ReleasesDGX agent

arXiv:2505.09368v2 Announce Type: replace Abstract: Standard benchmarks for optical flow, scene flow, and stereo vision algorithms generally focus on model accuracy rather than robustness to image cor

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

Thought Branches: Interpreting LLM Reasoning Requires Resampling

SafetyDGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

Towards Autonomous Mechanistic Reasoning in Virtual Cells

AgentsDGX agent

arXiv:2604.11661v1 Announce Type: cross Abstract: Large language models (LLMs) have recently gained significant attention as a promising approach to accelerate scientific discovery. However, their app

Zero-shot World Models Are Developmentally Efficient Learners

ResearchDGX agent

arXiv:2604.10333v1 Announce Type: new Abstract: Young children demonstrate early abilities to understand their physical world, estimating depth, motion, object coherence, interactions, and many other

13 Apr 2026

Another BRIXEL in the Wall: Towards Cheaper Dense Features

Local AiDGX agent

arXiv:2511.05168v2 Announce Type: replace Abstract: Vision foundation models achieve strong performance on both global and locally dense downstream tasks. Pretrained on large images, the recent DINOv3

Beyond Isolated Clients: Integrating Graph-Based Embeddings into Event Sequence Models

ResearchDGX agent

arXiv:2604.09085v1 Announce Type: cross Abstract: Large-scale digital platforms generate billions of timestamped user-item interactions (events) that are crucial for predicting user attributes in, e.g

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models

ApplicationsDGX agent

arXiv:2604.06161v2 Announce Type: replace-cross Abstract: Most digital videos are stored in 8-bit low dynamic range (LDR) formats, where much of the original high dynamic range (HDR) scene radiance is

Efficient Unlearning through Maximizing Relearning Convergence Delay

ResearchDGX agent

arXiv:2604.09391v1 Announce Type: cross Abstract: Machine unlearning poses challenges in removing mislabeled, contaminated, or problematic data from a pretrained model. Current unlearning approaches a

Gemma:26b thinking issue in openWebUI

Model ReleasesDGX agent

This r/ollama thread discusses user-reported issues with the Gemma 4 26B (a Mixture of Experts model) and its 'thinking' mode when used through Open WebUI. Key problems include the model getting stuck

← Previous
1…241242243244245…1018
Next →