AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,552 results
Agents

DriveFuture: Future-Aware Latent World Models for Autonomous Driving

DGX agent

arXiv:2605.09701v1 Announce Type: new Abstract: Existing latent world models for autonomous driving have opened a promising path toward future-aware driving intelligence. However, they typically treat

agentsarxiv-cs-cv
12 May 2026
Local Ai
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dystruct: Dynamically Structured Diffusion Language Model Decoding via Bayesian Inference

DGX agent

arXiv:2605.09820v1 Announce Type: new Abstract: Diffusion language models (DLMs) have recently emerged as a promising alternative to autoregressive models, primarily due to their ability to enable par

local-aiarxiv-cs-lg
12 May 2026
Research

Edit-Based Refinement for Parallel Masked Diffusion Language Models

DGX agent

arXiv:2605.09603v1 Announce Type: new Abstract: Masked diffusion language models enable parallel token generation and offer improved decoding efficiency over autoregressive models. However, their perf

researcharxiv-cs-cl
12 May 2026
Model Releases

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models

DGX agent

arXiv:2605.09773v1 Announce Type: cross Abstract: We use sparse autoencoder (SAE) feature steering to amplify Dark Triad personality traits (Machiavellianism, narcissism, and psychopathy) in Llama-3.3

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Failing Forward: Adaptive Failure-Informed Learning for Vision-Language-Action Models

DGX agent

arXiv:2605.08434v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising paradigm for scalable robotic manipulation, yet their reliance on success-only behavioral clonin

model-releasesarxiv-cs-ro
12 May 2026
Local Ai

FedGMI: Generative Model-Driven Federated Learning for Probabilistic Mixture Inference

DGX agent

arXiv:2605.08760v1 Announce Type: new Abstract: Federated Learning (FL) facilitates collaborative model training across decentralized clients while preserving data privacy by avoiding raw data exchang

local-aiarxiv-cs-lg
12 May 2026
Model Releases

Filtering Memorization from Parameter-Space in Diffusion Models

DGX agent

arXiv:2605.10439v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely used mechanism for customizing diffusion models, enabling users to inject new visual concepts or styles t

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models

DGX agent

arXiv:2603.16253v2 Announce Type: replace-cross Abstract: Vision-language process reward models (VL-PRMs) are increasingly used to score intermediate reasoning steps and rerank candidates under test-t

model-releasesarxiv-cs-ai
12 May 2026
Research

High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models

DGX agent

arXiv:2512.21815v2 Announce Type: replace Abstract: Vision-language models (VLMs) achieve remarkable performance but remain vulnerable to adversarial attacks. Entropy, as a measure of model uncertaint

researcharxiv-cs-cv
12 May 2026
Model Releases

Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models

DGX agent

arXiv:2605.08841v1 Announce Type: new Abstract: Vision-Language Models (VLMs) exhibit systematic bias toward visual illusions, recalling memorized facts rather than perceiving actual visual difference

model-releasesarxiv-cs-cv
12 May 2026
Research

Learngene Search Across Multiple Datasets for Building Variable-Sized Models

DGX agent

arXiv:2605.08209v1 Announce Type: new Abstract: Deep learning methods are widely used under diverse resource constraints, resulting in models of varying sizes, such as the Vision Transformer (ViT) ser

researcharxiv-cs-lg
12 May 2026
Model Releases

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models

DGX agent

arXiv:2510.08592v3 Announce Type: replace-cross Abstract: Test-Time Scaling (TTS) improves LLM reasoning by exploring multiple candidate responses and then operating over this set to find the best out

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LiteParse is the best open-source, model-free document parser for AI agents. Run it over over 50+ document types, and it will parse dense pa…

DGX agent

LiteParse is the best open-source, model-free document parser for AI agents. Run it over over 50+ document types, and it will parse dense pages with complex text layouts and tables, and it will extrac

model-releasesjerry-liu--x
12 May 2026
Model Releases

LLiMba: Sardinian on a Single GPU -- Adapting a 3B Language Model to a Vanishing Romance Language

DGX agent

arXiv:2605.09015v1 Announce Type: new Abstract: Sardinian, a Romance language with roughly one million speakers, has minimal presence in modern NLP. Commercial services do not support it, and current

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Med-StepBench: A Hierarchical Reasoning Framework for Evaluating Hallucinations in Medical Vision-Language Models

DGX agent

arXiv:2605.10002v1 Announce Type: new Abstract: Large vision-language models (VLMs) demonstrate strong performance in medical image understanding, but frequently generate clinically plausible yet inco

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Meow-Omni 1: A Multimodal Large Language Model for Feline Ethology

DGX agent

arXiv:2605.09152v1 Announce Type: new Abstract: Deciphering animal intent is a fundamental challenge in computational ethology, largely because of semantic aliasing, the phenomenon where identical ext

model-releasesarxiv-cs-cl
12 May 2026
Safety

On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models

DGX agent

arXiv:2605.09606v1 Announce Type: cross Abstract: Recent advances in image-to-3D models have significantly improved the fidelity and accessibility of 3D content creation. Such a powerful reconstructio

safetyarxiv-cs-cv
12 May 2026
Model Releases

Pretraining large language models with MXFP4

DGX agent

arXiv:2605.09825v1 Announce Type: cross Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We a

model-releasesarxiv-cs-ai
12 May 2026
Safety

SAID: Safety-Aware Intent Defense via Prefix Probing for Large Language Models

DGX agent

arXiv:2510.20129v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attacks, where adversarially crafted prompts induce policy-violating responses des

safetyarxiv-cs-ai
12 May 2026
Research

SLayerGen: a Crystal Generative Model for all Space and Layer Groups

DGX agent

arXiv:2605.08262v1 Announce Type: cross Abstract: Crystal generative models have shown rapid progress for accelerating the discovery of bulk, periodic materials. However, many material systems such as

researcharxiv-cs-ai
12 May 2026
Research

SlimQwen: Exploring the Pruning and Distillation in Large MoE Model Pre-training

DGX agent

arXiv:2605.08738v1 Announce Type: cross Abstract: Structured pruning and knowledge distillation (KD) are typical techniques for compressing large language models, but it remains unclear how they shoul

researcharxiv-cs-ai
12 May 2026
Model Releases

This NVIDIA remains the strongest platform for large-model inference at scale. Prefill/decode disaggregation, Blackwell-native quantization,…

DGX agent

This NVIDIA remains the strongest platform for large-model inference at scale. Prefill/decode disaggregation, Blackwell-native quantization, custom kernels, and rack-scale NVLink turn GB200 into faste

model-releasesperplexity--x
12 May 2026
Model Releases

Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition

DGX agent

arXiv:2605.10772v1 Announce Type: cross Abstract: Large language-vision models (LLVM), such as OpenAI's ChatGPT and GPT-4, have gained prominence as powerful tools for analyzing text and imagery. The

model-releasesarxiv-cs-ai
12 May 2026
Research

Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models

DGX agent

arXiv:2605.09005v1 Announce Type: cross Abstract: Vision-Language-Action models (VLAs) support generalist robotic control by enabling end-to-end decision policies directly from multi-modal inputs. As

researcharxiv-cs-ai
12 May 2026
Model Releases

Towards Compact Sign Language Translation: Frame Rate and Model Size Trade-offs

DGX agent

arXiv:2605.09554v1 Announce Type: new Abstract: Sign Language Translation (SLT) converts sign language videos into spoken-language text, bridging communication between Deaf and hearing communities. Cu

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models

DGX agent

arXiv:2605.08974v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have advanced video understanding, they remain highly prone to hallucinations in dynamic scenes. We arg

model-releasesarxiv-cs-ai
12 May 2026
Research

Uncovering Intra-expert Activation Sparsity for Efficient Mixture-of-Expert Model Execution

DGX agent

arXiv:2605.08575v1 Announce Type: cross Abstract: Mixture of Experts (MoE) architecture has become the standard for state-of-the-art large language models, owing to its computational efficiency throug

researcharxiv-cs-ai
12 May 2026
Model Releases

Understanding Asynchronous Inference Methods for Vision-Language-Action Models

DGX agent

arXiv:2605.08168v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising path to generalist robot control, but their inference latency causes observation staleness when

model-releasesarxiv-cs-ai
12 May 2026
Safety

Unsupervised Process Reward Models

DGX agent

arXiv:2605.10158v1 Announce Type: new Abstract: Process Reward Models (PRMs) are a powerful mechanism for steering large language model reasoning by providing fine-grained, step-level supervision. How

safetyarxiv-cs-lg
12 May 2026
Model Releases

Amortized Multi-Objective Optimization Across Tasks with Generative Solution Modeling

DGX agent

arXiv:2511.09598v5 Announce Type: replace Abstract: Many real-world applications require solving families of expensive multi-objective optimization problems~(EMOPs) under varying operational condition

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

BGM-IV: an AI-powered Bayesian generative modeling approach for instrumental variable analysis

DGX agent

arXiv:2605.07029v1 Announce Type: cross Abstract: Instrumental-variable (IV) regression enables causal estimation under endogeneity, but modern IV problems often involve nonlinear structural effects a

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

DT-PBO: an Interpretable Tree-based Surrogate Model for Preferential Bayesian Optimization

DGX agent

arXiv:2512.14263v2 Announce Type: replace-cross Abstract: Preferential Bayesian Optimization (PBO) aims to find a decision-maker's most preferred solution in as few pairwise comparisons as possible. E

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Efficient Data Selection for Multimodal Models via Incremental Optimization Utility

DGX agent

arXiv:2605.07488v1 Announce Type: new Abstract: The scaling of Large Multimodal Models (LMMs) is constrained by the quality-quantity trade-off inherent in synthetic data. Previous approaches, such as

model-releasesarxiv-cs-ai
11 May 2026
Safety

Flow-OPD: On-Policy Distillation for Flow Matching Models

DGX agent

arXiv:2605.08063v1 Announce Type: cross Abstract: Existing Flow Matching (FM) text-to-image models suffer from two critical bottlenecks under multi-task alignment: the reward sparsity induced by scala

safetyarxiv-cs-ai
11 May 2026
Model Releases

Frequency-Aware Model Parameter Explorer: A new attribution method for improving explainability

DGX agent

arXiv:2510.03245v2 Announce Type: replace-cross Abstract: State-of-the-art attribution methods rely on adversarial sample generation that applies an all-pass filter across the frequency spectrum, disc

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Introducing Daybreak: frontier AI for cyber defenders. Daybreak brings together the most capable OpenAI models, Codex, and our security part…

DGX agent

Introducing Daybreak: frontier AI for cyber defenders. Daybreak brings together the most capable OpenAI models, Codex, and our security partners to accelerate cyber defense and continuously secure sof

model-releasesopenai--x
11 May 2026
Applications

It is one reason why I think the push for smaller, local models is more complicated than people think. If you want good answers, especially …

DGX agent

It is one reason why I think the push for smaller, local models is more complicated than people think. If you want good answers, especially good answers to unexpected problems, frontier models will ge

applicationsethan-mollick--x
11 May 2026
Model Releases

MicroBi-ConvLSTM: An Ultra-Lightweight Efficient Model for Human Activity Recognition on Resource Constrained Devices

DGX agent

arXiv:2602.06523v2 Announce Type: replace Abstract: Human Activity Recognition (HAR) on resource constrained wearables requires models that balance accuracy against strict memory and computational bud

model-releasesarxiv-cs-cv
11 May 2026
Safety

NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models

DGX agent

arXiv:2605.07794v1 Announce Type: new Abstract: World Action Models (WAMs) are an emerging family of policies that tie robot action generation to future-observation modeling. In this work, we focus on

safetyarxiv-cs-ro
11 May 2026
Research

Normalizing Trajectory Models

DGX agent

arXiv:2605.08078v1 Announce Type: new Abstract: Diffusion-based models decompose sampling into many small Gaussian denoising steps -- an assumption that breaks down when generation is compressed to a

researcharxiv-cs-cv
11 May 2026
Applications

Our research, as well as that of other researchers, shows better prompting techniques help a lot, but model training is still a huge limitin…

DGX agent

Ethan Mollick discusses research findings showing that while improved prompting techniques provide significant benefits for AI model performance, the underlying model training remains the primary limi

applicationsethan-mollick--x
11 May 2026
Model Releases

Outlier Smoothing with Closed-Form Rotations for W4A4 Large Language Model Quantization

DGX agent

arXiv:2511.22316v2 Announce Type: replace Abstract: Large Language Models (LLMs) quantization facilitates deploying LLMs in resource-limited settings, but existing methods that combine incompatible gr

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

PathPainter: Transferring the Generalization Ability of Image Generation Models to Embodied Navigation

DGX agent

arXiv:2605.07496v1 Announce Type: new Abstract: Bird's-eye-view (BEV) images have been widely demonstrated to provide valuable prior information for navigation. Given the global information provided b

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models

DGX agent

arXiv:2605.07574v1 Announce Type: new Abstract: Mainstream vision-language models (VLMs) fundamentally struggle with severe optical ambiguities, such as reflections and transparent objects, due to the

model-releasesarxiv-cs-cv
11 May 2026
Research

ProteinJEPA: Latent prediction complements protein language models

DGX agent

arXiv:2605.07554v1 Announce Type: cross Abstract: Protein language models are trained primarily with masked language modeling (MLM), which predicts amino-acid identities at masked positions. We ask wh

researcharxiv-cs-ai
11 May 2026
Safety

Reflections and New Directions for Human-Centered Large Language Models

DGX agent

arXiv:2605.06901v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping the private and professional lives of users, with numerous applications in business, education, fi

safetyarxiv-cs-cl
11 May 2026
Safety

Stabilized neural Hamilton--Jacobi--Bellman solvers: Error analysis and applications in model-based reinforcement learning

DGX agent

arXiv:2605.07116v1 Announce Type: cross Abstract: Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is

safetyarxiv-cs-ai
11 May 2026
Model Releases

Structured Prototype-Guided Adaptation for EEG Foundation Models

DGX agent

arXiv:2602.17251v2 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (EFMs) have shown strong potential for transferable representation learning, yet their adaptation in

model-releasesarxiv-cs-lg
11 May 2026
← Previous
1…143144145146147…1262
Next →