AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Model Releases

Peer-Preservation in Frontier Models

DGX agent

arXiv:2604.19784v1 Announce Type: cross Abstract: Recently, it has been found that frontier AI models can resist their own shutdown, a behavior known as self-preservation. We extend this concept to th

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

SpanVLA: Efficient Action Bridging and Learning from Negative-Recovery Samples for Vision-Language-Action Model

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.19710v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising autonomous driving paradigm for leveraging world knowledge and reasoning capabilities, especially

safetyarxiv-cs-cv
22 Apr 2026
Model Releases

Interpolating Discrete Diffusion Models with Controllable Resampling

DGX agent

arXiv:2604.17310v1 Announce Type: new Abstract: Discrete diffusion models form a powerful class of generative models across diverse domains, including text and graphs. However, existing approaches fac

model-releasesarxiv-cs-lg
21 Apr 2026
Research

P-Check: Advancing Personalized Reward Model via Learning to Generate Dynamic Checklist

DGX agent

arXiv:2601.02986v2 Announce Type: replace Abstract: Recent approaches in personalized reward modeling have primarily focused on leveraging user interaction history to align model judgments with indivi

researcharxiv-cs-cl
21 Apr 2026
Model Releases

In Context Learning and Reasoning for Symbolic Regression with Large Language Models

DGX agent

arXiv:2410.17448v3 Announce Type: replace Abstract: Large Language Models (LLMs) are transformer-based machine learning models that have shown remarkable performance in tasks for which they were not e

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Three Roles, One Model: Role Orchestration at Inference Time to Close the Performance Gap Between Small and Large Agents

DGX agent

arXiv:2604.11465v1 Announce Type: new Abstract: Large language model (LLM) agents show promise on realistic tool-use tasks, but deploying capable agents on modest hardware remains challenging. We stud

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HiFloat4 Format for Language Model Pre-training on Ascend NPUs

DGX agent

arXiv:2604.08826v1 Announce Type: cross Abstract: Large foundation models have become central to modern machine learning, with performance scaling predictably with model size and data. However, traini

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Demystifying Adversarial Robustness in Diffusion Models: Compression, Randomness, and Geometry

DGX agent

arXiv:2505.22839v2 Announce Type: replace-cross Abstract: Recent studies suggest that diffusion models significantly improve the empirical adversarial robustness of deep neural network models. While i

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Can Open-Weight Models Compete on Financial Text Comprehension?

DGX agent

arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier models in recent months. Yet their reliability

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Benchmarking and Reasoning Distillation of Large Language Models for Feedback Controller Design in Complex Dynamical Systems

DGX agent

arXiv:2608.07004v1 Announce Type: new Abstract: Although remarkable capabilities have been demonstrated by Large Language Models (LLMs) across scientific domains, feedback controller design remains un

model-releasesarxiv-cs-ro
10 Aug 2026
Tutorials

Intrinsic-Hybrid Latent Diffusion Models for Generative Modeling on Unknown Manifolds

DGX agent

arXiv:2608.04827v1 Announce Type: cross Abstract: We introduce the Intrinsic Hybrid Latent Diffusion Model (ILDM), a generative framework that integrates probabilistic dimensionality reduction with ge

tutorialsarxiv-cs-lg
6 Aug 2026
Safety

A Security-Oriented Lifecycle Model for Large Language Model Systems

DGX agent

arXiv:2608.03626v1 Announce Type: cross Abstract: Large language models are being integrated into critical infrastructure and enterprise workflows at unprecedented scale,yet the lifecycle frameworks g

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis

DGX agent

arXiv:2608.00011v1 Announce Type: new Abstract: Current text-to-speech systems face a trade-off: autoregres- sive codec language models produce highly intelligible speech but require large-scale model

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

WorldExam: Benchmarking World Models from Apparent Appearance to Inherent Reactivity

DGX agent

arXiv:2608.02603v1 Announce Type: new Abstract: Controllable video generation models are increasingly being developed as world models. Accordingly, evaluating them in this role extends beyond the appa

model-releasesarxiv-cs-cv
4 Aug 2026
Research

A Lightweight Foundation Model for Collider Physics with Multi-Domain Adaptation

DGX agent

arXiv:2607.27501v1 Announce Type: new Abstract: We present a lightweight approach to foundation modeling (extbf{NEXUS}) that leverages pre-trained learning from collider physics data towards out-of-do

researcharxiv-cs-lg
31 Jul 2026
Model Releases

ECG-InterpBench: Benchmarking the Interpretability of ECG Foundation Models with Matched-Scale Sparse Autoencoders

DGX agent

arXiv:2607.27404v1 Announce Type: new Abstract: Existing benchmarks for electrocardiogram foundation models primarily evaluate downstream predictive performance, providing limited insight into whether

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation

DGX agent

arXiv:2607.27372v1 Announce Type: cross Abstract: The deep learning revolution, kicked off by AlexNet, taught us that end-to-end training beats decomposing a problem into hand-designed stages. Generat

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Metis: Memory Foundation Model

DGX agent

arXiv:2607.26760v1 Announce Type: new Abstract: Recent advances in AI agents have increasingly internalized native capabilities into their underlying foundation models, giving rise to multimodal found

agentsarxiv-cs-cl
30 Jul 2026
Applications

Memory Layer: Train the In-Model Cache for Recommendation Models

DGX agent

arXiv:2607.25110v1 Announce Type: cross Abstract: Early ranking stages in recommendation systems precompute item embeddings and cache them in-model for scoring within strict latency constraints. Becau

applicationsarxiv-cs-lg
29 Jul 2026
Model Releases

A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever

DGX agent

arXiv:2607.23806v1 Announce Type: cross Abstract: Improving a language model today means retraining it: enormous compute, a new opaque model each cycle, non-deterministic output. We take the opposite

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond Scale and Generation: Understanding Language Model-based Entity Matching

DGX agent

arXiv:2607.24688v1 Announce Type: cross Abstract: Entity matching identifies records that refer to the same real-world entity. Language models can be adapted to this task through bi-encoder, cross-enc

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Choosing a Text Embedding Model: A Practical Benchmarking and Decision Framework

DGX agent

arXiv:2607.23507v1 Announce Type: cross Abstract: Choosing the right text embedding model is one of the most consequential -- and most frequently under-examined -- decisions in building a retrieval or

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference

DGX agent

arXiv:2607.24585v1 Announce Type: new Abstract: We present ELMOD - Efficient Language Model for On-Device Deployment - a compact (2.7B) German language model designed for efficient inference on resour

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

UNIFUSION: Adapting Autoregressive Language Models into Discrete Diffusion under a Unified Reverse-Rate Objective

DGX agent

arXiv:2607.24507v1 Announce Type: cross Abstract: Existing methods mainly adapt pretrained autoregressive (AR) language models to masked diffusion, whereas we directly adapt them to uniform-noise diff

model-releasesarxiv-cs-ai
28 Jul 2026
Research

From Score Approximation to Distribution Approximation in Score-Based Diffusion Models

DGX agent

arXiv:2607.22199v1 Announce Type: new Abstract: Score-based diffusion models have achieved remarkable empirical success in generative modeling, yet their approximation-theoretic foundations remain inc

researcharxiv-cs-lg
27 Jul 2026
Model Releases

Improving Large Vision-Language Models' Understanding for Flow Field Data

DGX agent

arXiv:2507.18311v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have shown impressive capabilities across a range of tasks that integrate visual and textual understanding, suc

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Leaky Language Models: Stealing Architecture and Inference Optimizations via Per-Token Timing

DGX agent

arXiv:2607.20723v1 Announce Type: cross Abstract: This work presents LeakyLMs, a set of attacks that leak proprietary model, architecture, and deployment information from production language models. L

model-releasesarxiv-cs-lg
24 Jul 2026
Research

User-Centric Modeling of Transactional Sequences with Explainable State Space Models

DGX agent

arXiv:2607.20228v1 Announce Type: new Abstract: We propose a hybrid approach for user-centric modeling of transactional event sequences that combines contrastive representation learning (CoLES) with S

researcharxiv-cs-lg
23 Jul 2026
Model Releases

Agora: Collective and Permissionless Internet-Scale Pretraining of Large Language Models

DGX agent

arXiv:2607.13332v1 Announce Type: new Abstract: Training large language models at the multi-billion to trillion parameter scale is confined to datacenters, where data-parallel (DP) and model-parallel

model-releasesarxiv-cs-lg
16 Jul 2026
Research

A Threshold Exceedance Framework for CBRN Uplift Evaluation in Frontier Language Models

DGX agent

arXiv:2607.12200v1 Announce Type: new Abstract: As frontier language models advance, policymakers and model developers need methods for assessing whether model access materially increases a non-expert

researcharxiv-cs-ai
15 Jul 2026
Model Releases

APPLV: Adaptive Planner Parameter Learning from Vision-Language-Action Model

DGX agent

arXiv:2603.08862v2 Announce Type: replace-cross Abstract: Autonomous navigation in highly constrained environments remains challenging for mobile robots. Classical navigation approaches offer safety a

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

Scalable and Trustworthy Earth Observation Foundation Models

DGX agent

arXiv:2607.07758v1 Announce Type: new Abstract: Foundation models (FMs) have transformed machine learning from isolated task-specific model development toward general-purpose models pretrained on broa

model-releasesarxiv-cs-lg
10 Jul 2026
Research

Understanding Layer Patching in Model Size Interpolation

DGX agent

arXiv:2607.08170v1 Announce Type: new Abstract: Zero-shot model size interpolation aims to create new models of intermediate target sizes by combining existing models without additional training. Rece

researcharxiv-cs-lg
10 Jul 2026
Model Releases

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning

DGX agent

arXiv:2607.07690v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards (e.g. GRPO) is the engine behind today's reasoning models, yet it grades only the final answer. On hard

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Foundation Models for Automatic CAD Generation

DGX agent

arXiv:2607.05573v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) and Vision-Language Models (VLMs) enable the automatic generation of parametric 3D designs from natural-

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

The yes-no bias of large language models reflects answer order and wording, not shifts in moral judgment

DGX agent

arXiv:2607.05552v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly issue judgments read as binary verdicts, and a growing literature reports such judgments shifting under logi

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

X-FEMR: A Token-level Explainable Approach for Electronic Health Records Foundation Models using Transformer-based Models

DGX agent

arXiv:2607.06163v1 Announce Type: cross Abstract: Foundation Models for Electronic Health Records (FEMRs) are pretrained on large-scale structured patient data, enabling them to convert longitudinal p

safetyarxiv-cs-ai
8 Jul 2026
Research

CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space

DGX agent

arXiv:2512.08029v3 Announce Type: replace-cross Abstract: Clinical decision-making in oncology requires predicting dynamic disease evolution, a task current static AI predictors cannot perform. While

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Evaluating Intellectual Property Guardrails of Generative Image Models: A Technical Report

DGX agent

arXiv:2607.02582v1 Announce Type: new Abstract: Generative image models are capable of producing images that bear a strong resemblance to, or replicate, recognizable intellectual property (IP). In thi

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

EduArt: An educational-level benchmark for evaluating art history knowledge in large language models

DGX agent

arXiv:2607.02007v1 Announce Type: new Abstract: Large language models now score near ceiling on general benchmarks, but these aggregate measures reveal little about how models behave within single dis

model-releasesarxiv-cs-cl
3 Jul 2026
Safety

YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models

DGX agent

arXiv:2601.15588v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, safety guardrails are required to go beyond coarse-grained fil

safetyarxiv-cs-cl
3 Jul 2026
Model Releases

Beyond IID: How General Are Tabular Foundation Models, Really?

DGX agent

arXiv:2606.30410v1 Announce Type: cross Abstract: Foundation models for predictive machine learning on tabular data have recently gained significant traction in academia and industry. Research communi

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencoders

DGX agent

arXiv:2507.23220v2 Announce Type: replace Abstract: Traditional topic models are effective at uncovering latent themes in large text collections. However, due to their reliance on bag-of-words represe

researcharxiv-cs-cl
30 Jun 2026
Safety

In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis from Language Models to Physics

DGX agent

arXiv:2606.26981v1 Announce Type: cross Abstract: Synthesizing human motion from textual descriptions is essential for immersive digital applications, yet existing methods face a persistent trade-off

safetyarxiv-cs-ai
26 Jun 2026
Tutorials

Multifidelity-Augmented Gaussian Process Inputs for Surrogate Modeling from Scarce Data

DGX agent

arXiv:2603.22050v2 Announce Type: replace-cross Abstract: Supervised machine learning describes the practice of fitting a parameterized model to labeled input-output data. Supervised machine learning

tutorialsarxiv-cs-lg
25 Jun 2026
Model Releases

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks

DGX agent

arXiv:2606.24162v1 Announce Type: new Abstract: Foundation models have been increasingly applied to behavioral science domains such as psychology, sociology, and economics. While these models show pro

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

World Value Models for Robotic Manipulation

DGX agent

arXiv:2606.24742v1 Announce Type: new Abstract: Generalist value models play a pivotal role in scaling robotic policy learning from large-scale, mixed-quality data. Mathematically, accurate value esti

model-releasesarxiv-cs-ro
24 Jun 2026
Safety

A Watermark for Vision-Language-Action and World Action Models

DGX agent

arXiv:2606.23574v1 Announce Type: cross Abstract: Vision-language-action (VLA) models and world-action models (WAM) are the generative models now driving general-purpose robot control, turning raw cam

safetyarxiv-cs-ro
23 Jun 2026
← Previous
1…678910…1012
Next →