AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,552 results
Model Releases

any time a model router company drops data, its worth browsing. here we learn that gemini leads in education and personal assistants (?!), a…

DGX agent

any time a model router company drops data, its worth browsing. here we learn that gemini leads in education and personal assistants (?!), ant leads in vibecoding and koding and back office (?!), and

model-releasesswyx--x
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

CADDesigner: Conceptual CAD Model Generation with a General-Purpose Agent

DGX agent

arXiv:2508.01031v5 Announce Type: replace Abstract: Computer-Aided Design (CAD) is widely used for conceptual design and parametric 3D modeling, but typically requires a high level of expertise from d

agentsarxiv-cs-ai
14 May 2026
Safety

ChatSR: Multimodal Large Language Models for Scientific Formula Discovery

DGX agent

arXiv:2406.05410v3 Announce Type: replace Abstract: Current multimodal large language models (MLLMs) are mainly focused on the understanding and processing of perceptual modalities such as images and

safetyarxiv-cs-ai
14 May 2026
Model Releases

Compact Latent Manifold Translation: A Parameter-Efficient Foundation Model for Cross-Modal and Cross-Frequency Physiological Signal Synthesis

DGX agent

arXiv:2605.13248v1 Announce Type: cross Abstract: The analysis of physiological time series, such as electrocardiograms (ECG) and photoplethysmograms (PPG), is persistently hindered by modality and fr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Connecting the Dots: A Machine Learning Ready Dataset for Ionospheric Forecasting Models

DGX agent

arXiv:2511.15743v2 Announce Type: replace Abstract: Operational forecasting of the ionosphere remains a critical space weather challenge due to sparse observations, complex coupling across geospatial

model-releasesarxiv-cs-lg
14 May 2026
Research

Continual Learning with Multilingual Foundation Model

DGX agent

arXiv:2605.13415v1 Announce Type: cross Abstract: This paper presents a multi-stage framework for detecting reclaimed slurs in multilingual social media discourse. It addresses the challenge of identi

researcharxiv-cs-ai
14 May 2026
Safety

Diffusion Model's Generalization Can Be Characterized by Inductive Biases toward a Data-Dependent Ridge Manifold

DGX agent

arXiv:2602.06021v2 Announce Type: replace-cross Abstract: We study a data-dependent notion of diffusion-model generalization: when a model does not memorize the training set, where do its generated sa

safetyarxiv-cs-lg
14 May 2026
Safety

DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense

DGX agent

arXiv:2605.13115v1 Announce Type: cross Abstract: Diffusion models depend on pseudo-random number generators (PRNGs) for latent noise sampling. We present DiffusionHijack, a supply-chain backdoor atta

safetyarxiv-cs-lg
14 May 2026
Model Releases

DistractMIA: Black-Box Membership Inference on Vision-Language Models via Semantic Distraction

DGX agent

arXiv:2605.12574v1 Announce Type: cross Abstract: Vision-language models (VLMs) are trained on large-scale image-text corpora that may contain private, copyrighted, or otherwise sensitive data, motiva

model-releasesarxiv-cs-ai
14 May 2026
Applications

It's not the Language Model, it's the Tool: Deterministic Mediation for Scientific Workflows

DGX agent

arXiv:2605.13245v1 Announce Type: new Abstract: Language models can produce convincing scientific analyses, but repeated generations on the same data do not guarantee the same result. A researcher may

applicationsarxiv-cs-ai
14 May 2026
Local Ai

KAST-BAR: Knowledge-Anchored Semantically-Dynamic Topology Brain Autoregressive Modeling for Universal Neural Interpretation

DGX agent

arXiv:2605.13133v1 Announce Type: new Abstract: While EEG foundation models have shown significant potential in universal neural decoding across tasks, their advancement remains constrained by the ina

local-aiarxiv-cs-lg
14 May 2026
Research

Latent-Augmented Discrete Diffusion Models

DGX agent

arXiv:2510.18114v3 Announce Type: replace-cross Abstract: Discrete diffusion models have emerged as a powerful class of models and a promising route to fast language generation, but practical implemen

researcharxiv-cs-ai
14 May 2026
Model Releases

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in …

DGX agent

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in 2026. And like we all know, some 10x leaps can't survive a c

model-releasesfireworks-ai--x
14 May 2026
Local Ai

LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters

DGX agent

arXiv:2605.13163v1 Announce Type: cross Abstract: Foundation models and low-rank adapters enable efficient on-device generative AI but raise risks such as intellectual property leakage and model recov

local-aiarxiv-cs-cv
14 May 2026
Agents

Mechanism Plausibility in Generative Agent-Based Modeling

DGX agent

arXiv:2605.12824v1 Announce Type: cross Abstract: Large language models (LLMs) can generate high-level diverse phenomena without explicitly programmed rules. This capability has led to their adoption

agentsarxiv-cs-ai
14 May 2026
Research

MIDST Challenge at SaTML 2025: Membership Inference over Diffusion-models-based Synthetic Tabular data

DGX agent

arXiv:2603.19185v2 Announce Type: replace Abstract: Synthetic data is often perceived as a silver-bullet solution to data anonymization and privacy-preserving data publishing. Drawn from generative mo

researcharxiv-cs-lg
14 May 2026
Research

QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling

DGX agent

arXiv:2605.13833v1 Announce Type: cross Abstract: Modeling long-range dependencies in sequential data remains a central challenge in machine learning. Transformers address this challenge through atten

researcharxiv-cs-cv
14 May 2026
Research

The Expressivity Boundary of Probabilistic Circuits: A Comparison with Large Language Models

DGX agent

arXiv:2605.12940v1 Announce Type: cross Abstract: Probabilistic Circuits (PCs) are deep generative models that support exact and efficient probabilistic inference. Yet in autoregressive language model

researcharxiv-cs-ai
14 May 2026
Research

The Score-Difference Flow for Implicit Generative Modeling

DGX agent

arXiv:2304.12906v4 Announce Type: replace Abstract: Implicit generative modeling (IGM) aims to produce samples of synthetic data matching the characteristics of a target data distribution. Recent work

researcharxiv-cs-lg
14 May 2026
Research

VIP-COP: Context Optimization for Tabular Foundation Models

DGX agent

arXiv:2605.12904v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have emerged as a powerful paradigm for in-context learning on structured data, enabling direct prediction on new tabul

researcharxiv-cs-lg
14 May 2026
Safety

What properties of reasoning supervision are associated with improved downstream model quality?

DGX agent

arXiv:2605.13290v1 Announce Type: new Abstract: Validating training data for reasoning models typically requires expensive trial-and-error fine-tuning cycles. In this work, we investigate whether the

safetyarxiv-cs-ai
14 May 2026
Model Releases

A Study on Hidden Layer Distillation for Large Language Model Pre-Training

DGX agent

arXiv:2605.11513v1 Announce Type: new Abstract: Knowledge Distillation (KD) is a critical tool for training Large Language Models (LLMs), yet the majority of research focuses on approaches that rely s

model-releasesarxiv-cs-cl
13 May 2026
Research

Active inference as a unified model of collision avoidance behavior in human drivers

DGX agent

arXiv:2506.02215v5 Announce Type: replace Abstract: Collision avoidance -- involving a rapid threat detection and quick execution of the appropriate evasive maneuver -- is a critical aspect of driving

researcharxiv-cs-ro
13 May 2026
Model Releases

ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder

DGX agent

arXiv:2605.11091v1 Announce Type: new Abstract: Automated ASD screening tools remain limited by single-architecture evaluations, axis-restricted assessment, and near-exclusive focus on adult cohorts,

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models

DGX agent

arXiv:2605.11726v1 Announce Type: new Abstract: Recently, reinforcement learning (RL) has been widely applied during post-training for diffusion large language models (dLLMs) to enhance reasoning with

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Built a local coding harness powered by Gemma 4. It runs locally, connects to my model backend, starts coding sessions, streams responses, a…

DGX agent

Built a local coding harness powered by Gemma 4. It runs locally, connects to my model backend, starts coding sessions, streams responses, and uses tools through a CLI-style workflow. Still early, but

model-releasesollama--x
13 May 2026
Safety

Clarity: The Flexibility-Interpretability Trade-Off in Sparsity-aware Concept Bottleneck Models

DGX agent

arXiv:2601.21944v2 Announce Type: replace Abstract: The widespread adoption of deep learning models in computer vision has intensified concerns about interpretability. Despite strong performance, thes

safetyarxiv-cs-lg
13 May 2026
Safety

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion

DGX agent

arXiv:2505.18780v3 Announce Type: replace-cross Abstract: Achieving versatile humanoid locomotion with a single policy presents a critical scalability challenge. Prevailing methods often rely on disti

safetyarxiv-cs-lg
13 May 2026
Safety

EHR-RAGp: Retrieval-Augmented Prototype-Guided Foundation Model for Electronic Health Records

DGX agent

arXiv:2605.12335v1 Announce Type: cross Abstract: Electronic Health Records (EHR) contain rich longitudinal patient information and are widely used in predictive modeling applications. However, effect

safetyarxiv-cs-lg
13 May 2026
Research

Express Your Doubts -- Probabilistic World Modeling Should not be Based on Token logprobs

DGX agent

arXiv:2505.02072v2 Announce Type: replace Abstract: Language modeling has shifted in recent years from a distribution over strings to prediction models with textual inputs and outputs for general-purp

researcharxiv-cs-cl
13 May 2026
Safety

Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2605.12374v1 Announce Type: new Abstract: Visual latent reasoning lets a multimodal large language model (MLLM) create intermediate visual evidence as continuous tokens, avoiding external tools

safetyarxiv-cs-cv
13 May 2026
Model Releases

Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems

DGX agent

arXiv:2603.28561v2 Announce Type: replace Abstract: The growing deployment of small Unmanned Aerial Systems (sUASs) in low-altitude airspaces has increased the need for reliable tactical deconfliction

model-releasesarxiv-cs-ro
13 May 2026
Research

Molecular Design beyond Training Data with Novel Extended Objective Functionals of Generative AI Models Driven by Quantum Annealing Computer

DGX agent

arXiv:2602.15451v3 Announce Type: replace-cross Abstract: Deep generative modeling to stochastically design small molecules is an emerging technology for accelerating drug discovery and development. H

researcharxiv-cs-lg
13 May 2026
Safety

PriorZero: Bridging Language Priors and World Models for Decision Making

DGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

safetyarxiv-cs-lg
13 May 2026
Research

@SakanaAILabs @NVIDIAAI Sparser, Faster, Lighter Transformer Language Models https://arxiv.org/abs/2603.23198

DGX agent

This research paper from Sakana AI and NVIDIA explores techniques for creating more efficient transformer language models by reducing sparsity, computational requirements, and model size while maintai

researchdavid-ha--x
13 May 2026
Model Releases

See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model

DGX agent

arXiv:2605.11817v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown remarkable promise in robotics manipulation, yet their high computational cost hinders real-time deploy

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Sources: Mistral has been developing a cybersecurity-focused AI model and held discussions about it with European banks, which don't have access to Mythos (Bloomberg)

DGX agent

Bloomberg: Sources: Mistral has been developing a cybersecurity-focused AI model and held discussions about it with European banks, which don't have access to Mythos — French artificial intelligence s

model-releasestechmeme
13 May 2026
Safety

STRIDE: Training-Free Diversity Guidance via PCA-Directed Feature Perturbation in Single-Step Diffusion Models

DGX agent

arXiv:2605.11494v1 Announce Type: new Abstract: Distilled one-step (T=1) or few-step (Tleq4) diffusion models enable real-time image generation but often exhibit reduced sample diversity compared to t

safetyarxiv-cs-cv
13 May 2026
Local Ai

Structural Interpretations of Protein Language Model Representations via Differentiable Graph Partitioning

DGX agent

arXiv:2605.10985v1 Announce Type: new Abstract: Protein language models such as ESM-2 learn rich residue representations that achieve strong performance on protein function prediction, but their featu

local-aiarxiv-cs-lg
13 May 2026
Model Releases

Unlocking UML Class Diagram Understanding in Vision Language Models

DGX agent

arXiv:2605.11634v1 Announce Type: new Abstract: Although Vision Language Models (VLMs) have seen tremendous progress across all kinds of use cases, they still fall behind in answering questions regard

model-releasesarxiv-cs-cv
13 May 2026
Research

Vision-aligned Latent Reasoning for Multi-modal Large Language Model

DGX agent

arXiv:2602.04476v2 Announce Type: replace Abstract: Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems whi

researcharxiv-cs-cv
13 May 2026
Research

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

DGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

researcharxiv-cs-cl
13 May 2026
Model Releases

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

DGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

Attention Drift: What Autoregressive Speculative Decoding Models Learn

DGX agent

arXiv:2605.09992v1 Announce Type: cross Abstract: Speculative decoding accelerates LLM inference by drafting future tokens with a small model, but drafter models degrade sharply under template perturb

tutorialsarxiv-cs-ai
12 May 2026
Safety

Bi-CoG: Bi-Consistency-Guided Self-Training for Vision-Language Models

DGX agent

arXiv:2510.20477v2 Announce Type: replace Abstract: Exploiting unlabeled data through semi-supervised learning (SSL) or leveraging pre-trained models via fine-tuning are two prevailing paradigms for a

safetyarxiv-cs-lg
12 May 2026
Model Releases

C-CoT: Counterfactual Chain-of-Thought with Vision-Language Models for Safe Autonomous Driving

DGX agent

arXiv:2605.10744v1 Announce Type: new Abstract: Safety-critical planning in complex environments, particularly at urban intersections, remains a fundamental challenge for autonomous driving. Existing

model-releasesarxiv-cs-cv
12 May 2026
Research

DARE: Diffusion Language Model Activation Reuse for Efficient Inference

DGX agent

arXiv:2605.08134v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to auto-regressive (AR) models, offering greater expressive capacity a

researcharxiv-cs-ai
12 May 2026
Research

Do multimodal models imagine electric sheep?

DGX agent

arXiv:2605.09693v1 Announce Type: cross Abstract: Yes. We find that large multimodal models develop mental imagery when solving spatial puzzles, and they do imagine sheep when solving sheep puzzles. W

researcharxiv-cs-ai
12 May 2026
← Previous
1…142143144145146…1262
Next →