AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Prune, Update and Trim: Robust Structured Pruning for Large Language Models

DGX agent

arXiv:2605.18331v1 Announce Type: new Abstract: Large Language Models (LLMs) have experienced significant growth and development in recent years. However, performing inference on LLMs remains costly,

researcharxiv-cs-lg
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models

DGX agent

arXiv:2501.17549v2 Announce Type: replace Abstract: Graph-structured data plays a vital role in numerous domains, such as social networks, citation networks, commonsense reasoning graphs and knowledge

model-releasesarxiv-cs-cl
19 May 2026
Applications

Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift

DGX agent

arXiv:2605.16411v1 Announce Type: cross Abstract: Hallucination remains a fundamental challenge in vision-language models (VLMs), where autoregressive generation may produce linguistically plausible y

applicationsarxiv-cs-ai
19 May 2026
Safety

Retrieval and competition: how a protein foundation model starts a protein

DGX agent

arXiv:2605.16331v1 Announce Type: cross Abstract: Protein language models are increasingly used to guide experimental and clinical decisions, yet it is often unclear whether a confident prediction ref

safetyarxiv-cs-ai
19 May 2026
Model Releases

StableVLA: Towards Robust Vision-Language-Action Models without Extra Data

DGX agent

arXiv:2605.18287v1 Announce Type: new Abstract: It is infeasible to encompass all possible disturbances within the training dataset. This raises a critical question regarding the robustness of Vision-

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Structured Labeling Enables Faster Vision-Language Models for End-to-End Autonomous Driving

DGX agent

arXiv:2506.05442v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) offer a promising approach to end-to-end autonomous driving due to their human-like reasoning capabilities. Howe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling

DGX agent

arXiv:2605.17568v1 Announce Type: new Abstract: Multi-class event streams arise in numerous real-world applications, where uncovering structured, interpretable inter-event relationships, together with

model-releasesarxiv-cs-lg
19 May 2026
Hardware

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

DGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

hardwarearxiv-cs-ai
19 May 2026
Model Releases

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models

DGX agent

arXiv:2605.17577v1 Announce Type: new Abstract: Large-scale pre-trained Vision-Language models (VLMs), such as CLIP, exhibit strong zero-shot generalization, yet remain highly vulnerable to impercepti

model-releasesarxiv-cs-cv
19 May 2026
Safety

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving

DGX agent

arXiv:2602.22801v2 Announce Type: replace-cross Abstract: Diffusion models have become a popular choice for decision-making tasks in robotics, and more recently, are also being considered for solving

safetyarxiv-cs-ai
19 May 2026
Research

Unlocking the Potential of Diffusion Language Models through Template Infilling

DGX agent

arXiv:2510.13870v3 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) have emerged as a promising alternative to Autoregressive Language Models, yet their inference strategies rem

researcharxiv-cs-ai
19 May 2026
Model Releases

WOW-Seg: A Word-free Open World Segmentation Model

DGX agent

arXiv:2605.16903v1 Announce Type: new Abstract: Open world image segmentation aims to achieve precise segmentation and semantic understanding of targets within images by addressing the infinitely open

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Deterministic Event-Graph Substrates as World Models for Counterfactual Reasoning

DGX agent

arXiv:2605.15967v1 Announce Type: new Abstract: We study event-graph substrates: a class of world models that represent agent state as an append-only log of typed RDF triples and answer counterfactual

model-releasesarxiv-cs-ai
18 May 2026
Tutorials

DiscussLLM: Teaching Large Language Models When to Speak

DGX agent

arXiv:2508.18167v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in understanding and generating human-like text, yet they largely operate as

tutorialsarxiv-cs-cl
18 May 2026
Model Releases

Dynamic Chunking for Diffusion Language Models

DGX agent

arXiv:2605.15676v1 Announce Type: new Abstract: Block discrete diffusion language models factorize a sequence autoregressively over fixed-size positional blocks, decoupling within-block parallel denoi

model-releasesarxiv-cs-cl
18 May 2026
Research

Dynamics-Level Watermarking of Flow Matching Models with Random Codes

DGX agent

arXiv:2605.16239v1 Announce Type: new Abstract: We introduce a dynamics-level approach to watermarking generative models. Rather than embedding signals into model weights or outputs, we embed the wate

researcharxiv-cs-lg
18 May 2026
Research

Enabling Adversarial Robustness in AI Models through Kubeflow MLOps

DGX agent

arXiv:2605.15249v1 Announce Type: cross Abstract: AI models are increasingly deployed in cloud-native environments to support scalable and automated services. However, while platforms such as Kubernet

researcharxiv-cs-lg
18 May 2026
Model Releases

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

DGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Measuring Maximum Activations in Open Large Language Models

DGX agent

arXiv:2605.15572v1 Announce Type: new Abstract: The dynamic range of activations is a first-order constraint for low-bit quantization, activation scaling, and stable LLM inference. Prior work characte

model-releasesarxiv-cs-cl
18 May 2026
Local Ai

MIND: Decoupling Model-Induced Label Noise via Latent Manifold Disentanglement

DGX agent

arXiv:2605.16081v1 Announce Type: cross Abstract: The paradigm of learning from automatic annotations driven by pre-trained experts and Foundation Models dominates data-hungry applications. However, i

local-aiarxiv-cs-cv
18 May 2026
Safety

Offline Reinforcement Learning with Universal Horizon Models

DGX agent

arXiv:2605.15603v1 Announce Type: cross Abstract: Model-based reinforcement learning (RL) offers a compelling approach to offline RL by enabling value learning on imagined on-policy trajectories. Howe

safetyarxiv-cs-ai
18 May 2026
Research

Overfitting has a limitation: a model-independent generalization gap bound based on Renyi entropy

DGX agent

arXiv:2506.00182v3 Announce Type: replace-cross Abstract: Will further scaling up of machine learning models continue to bring success? A significant challenge in answering this question lies in under

researcharxiv-cs-lg
18 May 2026
Model Releases

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

DGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

model-releasesarxiv-cs-cl
18 May 2026
Safety

ReactiveGWM: Steering NPC in Reactive Game World Models

DGX agent

arXiv:2605.15256v1 Announce Type: new Abstract: Current game world models simulate environments from a subjective, player-centric perspective. However, by treating the Non-Player Character (NPC) merel

safetyarxiv-cs-cv
18 May 2026
Research

Reasoning Models Don't Just Think Longer, They Move Differently

DGX agent

arXiv:2605.15454v1 Announce Type: new Abstract: Reasoning-trained language models often spend more tokens on harder problems, but longer chains of thought do not show whether a model is merely computi

researcharxiv-cs-cl
18 May 2026
Local Ai

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers

DGX agent

arXiv:2505.12601v2 Announce Type: replace Abstract: As large language models (LLMs) grow in scale and specialization, routing--selecting the best model for a given input--has become essential for effi

local-aiarxiv-cs-lg
18 May 2026
Safety

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

DGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

safetyarxiv-cs-cv
18 May 2026
Research

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language

DGX agent

arXiv:2605.15607v1 Announce Type: new Abstract: Large language models (LLMs) achieve high pass rates on code generation benchmarks, yet whether they can transfer this ability to languages absent from

researcharxiv-cs-cl
18 May 2026
Model Releases

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

DGX agent

arXiv:2602.08324v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning successfully enhances the reasoning capabilities of Large Language Models (LLMs), yet it incurs substantial computa

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Action-Inspired Generative Models

DGX agent

arXiv:2605.14631v1 Announce Type: cross Abstract: We introduce Action-Inspired Generative Models (AGMs), a dual-network generative framework motivated by the observation that existing bridge-matching

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models

DGX agent

arXiv:2605.14897v1 Announce Type: cross Abstract: Despite many successful attempts at explaining Deep Reinforcement Learning policies using distillation, it remains difficult to balance the performanc

model-releasesarxiv-cs-ai
15 May 2026
Safety

EponaV2: Driving World Model with Comprehensive Future Reasoning

DGX agent

arXiv:2605.14696v1 Announce Type: new Abstract: Data scaling plays a pivotal role in the pursuit of general intelligence. However, the prevailing perception-planning paradigm in autonomous driving rel

safetyarxiv-cs-cv
15 May 2026
Safety

Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model

DGX agent

arXiv:2605.14950v1 Announce Type: new Abstract: Vision-Language-Action models have emerged as a promising paradigm for robotic manipulation by unifying perception, language grounding, and action gener

safetyarxiv-cs-cv
15 May 2026
Model Releases

FedStain: Modeling Higher-Order Stain Statistics for Federated Domain Generalization in Computational Pathology

DGX agent

arXiv:2605.14590v1 Announce Type: new Abstract: Robust whole-slide image (WSI) analysis under strict data-governance remains challenging due to substantial cross-institutional stain heterogeneity. Dom

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models

DGX agent

arXiv:2605.14906v1 Announce Type: new Abstract: Memory is essential for large vision-language models (LVLMs) to handle long, multimodal interactions, with two method directions providing this capabili

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Octopus: History-Free Gradient Orthogonalization for Continual Learning in Multimodal Large Language Models

DGX agent

arXiv:2605.14938v1 Announce Type: cross Abstract: Continual learning in multimodal large language models (MLLMs) aims to sequentially acquire knowledge while mitigating catastrophic forgetting, yet ex

model-releasesarxiv-cs-cv
15 May 2026
Safety

Quantitative Video World Model Evaluation for Geometric-Consistency

DGX agent

arXiv:2605.15185v1 Announce Type: cross Abstract: Generative video models are increasingly studied as implicit world models, yet evaluating whether they produce physically plausible 3D structure and m

safetyarxiv-cs-ai
15 May 2026
Research

RePack then Refine: Efficient Diffusion Transformer with Vision Foundation Model

DGX agent

arXiv:2512.12083v3 Announce Type: replace Abstract: Semantic-rich features from Vision Foundation Models (VFMs) have been leveraged to enhance Latent Diffusion Models (LDMs). However, raw VFM features

researcharxiv-cs-cv
15 May 2026
Safety

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models

DGX agent

arXiv:2512.21651v2 Announce Type: replace Abstract: Large Language Models (LLMs) deliver strong performance across a wide range of NLP tasks, but their massive sizes hinder deployment on resource-cons

safetyarxiv-cs-lg
15 May 2026
Applications

Scalable Krylov Subspace Methods for Generalized Mixed-Effects Models with Crossed Random Effects

DGX agent

arXiv:2505.09552v3 Announce Type: replace-cross Abstract: Mixed-effects models are widely used to model data with hierarchical grouping structures and high-cardinality categorical predictor variables.

applicationsarxiv-cs-lg
15 May 2026
Model Releases

SemaTune: Semantic-Aware Online OS Tuning with Large Language Models

DGX agent

arXiv:2605.15026v1 Announce Type: cross Abstract: Online OS tuning can improve long-running services, but existing controllers are poorly matched to live hosts. They treat scheduler, power, memory, an

model-releasesarxiv-cs-ai
15 May 2026
Research

Teaching Large Language Models When Not to Know: Learning Temporal Critique for Ex-Ante Reasoning

DGX agent

arXiv:2605.14636v1 Announce Type: new Abstract: Large language models (LLMs) often fail to reason under temporal cutoffs: when prompted to answer from the standpoint of an earlier time, they exploit k

researcharxiv-cs-ai
15 May 2026
Research

Uncertainty Quantification for Large Language Diffusion Models

DGX agent

arXiv:2605.14570v1 Announce Type: new Abstract: Large Language Diffusion Models (LLDMs) are emerging as an alternative to autoregressive models, offering faster inference through higher parallelism. S

researcharxiv-cs-cl
15 May 2026
Research

A Markov Categorical Framework for Language Modeling

DGX agent

arXiv:2507.19247v5 Announce Type: replace-cross Abstract: Autoregressive language models achieve remarkable performance, yet a unified theory explaining their internal mechanisms, how training shapes

researcharxiv-cs-ai
14 May 2026
Agents

CADDesigner: Conceptual CAD Model Generation with a General-Purpose Agent

DGX agent

arXiv:2508.01031v5 Announce Type: replace Abstract: Computer-Aided Design (CAD) is widely used for conceptual design and parametric 3D modeling, but typically requires a high level of expertise from d

agentsarxiv-cs-ai
14 May 2026
Safety

ChatSR: Multimodal Large Language Models for Scientific Formula Discovery

DGX agent

arXiv:2406.05410v3 Announce Type: replace Abstract: Current multimodal large language models (MLLMs) are mainly focused on the understanding and processing of perceptual modalities such as images and

safetyarxiv-cs-ai
14 May 2026
Model Releases

Compact Latent Manifold Translation: A Parameter-Efficient Foundation Model for Cross-Modal and Cross-Frequency Physiological Signal Synthesis

DGX agent

arXiv:2605.13248v1 Announce Type: cross Abstract: The analysis of physiological time series, such as electrocardiograms (ECG) and photoplethysmograms (PPG), is persistently hindered by modality and fr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Connecting the Dots: A Machine Learning Ready Dataset for Ionospheric Forecasting Models

DGX agent

arXiv:2511.15743v2 Announce Type: replace Abstract: Operational forecasting of the ionosphere remains a critical space weather challenge due to sparse observations, complex coupling across geospatial

model-releasesarxiv-cs-lg
14 May 2026
← Previous
1…111112113114115…1030
Next →