AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,404 results
Model Releases

Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement

DGX agent

arXiv:2602.03983v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for generalist robotic control. Built upon vision-language model (

model-releasesarxiv-cs-ro
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …

DGX agent

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open models. Some ideas for what comes next, May 2026 Gemini Flas

model-releasesclem-delangue--x
26 May 2026
Research

Nano World Models: A Minimalist Implementation of Future Video Prediction

DGX agent

arXiv:2605.23993v1 Announce Type: cross Abstract: World models have become a central paradigm for learning predictive simulators that support generation, planning, and decision-making. Yet, despite ra

researcharxiv-cs-ai
26 May 2026
Model Releases

One-for-All Model Initialization with Frequency-Domain Knowledge

DGX agent

arXiv:2603.07523v2 Announce Type: replace Abstract: Transferring knowledge by fine-tuning large-scale pre-trained networks has become a standard paradigm for downstream tasks, yet the knowledge of a p

model-releasesarxiv-cs-lg
26 May 2026
Agents

Practical Quantum CIM Empowerment via All-Domestic-Core Agentic Large Model

DGX agent

arXiv:2605.23934v1 Announce Type: new Abstract: Quantum computing devices are recognized as powerful tools for solving NP-complete problems. However, the intricacy of their modeling presents notable b

agentsarxiv-cs-ai
26 May 2026
Model Releases

Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers

DGX agent

arXiv:2605.25949v1 Announce Type: cross Abstract: Neural PDE solvers have followed the scaling trajectory of vision and language, with recent foundation models reaching billions of parameters. We argu

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

DGX agent

arXiv:2506.18543v2 Announce Type: replace-cross Abstract: The rapid proliferation of Large Language Models (LLMs) has heightened concerns regarding their exposure to jailbreak attacks, which craft adv

model-releasesarxiv-cs-ai
26 May 2026
Research

T2S-MPC: Time-Embedded Online Adaptive Model Predictive Control for Time-Varying Dynamics

DGX agent

arXiv:2605.24852v1 Announce Type: new Abstract: Recent advances in learning-based model predictive control (MPC) have leveraged neural networks for online model learning, achieving strong performance

researcharxiv-cs-lg
26 May 2026
Model Releases

The Model Is Not the Product: A Dual-Pillar Architecture for Local-First Psychological Coaching

DGX agent

arXiv:2605.24411v1 Announce Type: new Abstract: Existing language model applications struggle to meet the demand for emotionally oriented support, primarily due to their inability to maintain deep, pe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation

DGX agent

arXiv:2602.23916v2 Announce Type: replace-cross Abstract: The advent of large-scale self-supervised learning (SSL) has produced a vast zoo of medical foundation models. However, selecting optimal medi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Towards Large Model Feature Coding

DGX agent

arXiv:2605.24025v1 Announce Type: cross Abstract: Large models have delivered remarkable performance across a wide range of perception and generation tasks, yet practical deployment is increasingly co

model-releasesarxiv-cs-lg
26 May 2026
Research

TUBE: Tangent Upper Bound on Evidence for Discrete Diffusion Language Models

DGX agent

arXiv:2605.24292v1 Announce Type: new Abstract: Log-likelihood is a standard metric for evaluating generative models. Unfortunately, in contrast to autoregressive models (ARMs), discrete diffusion mod

researcharxiv-cs-lg
26 May 2026
Model Releases

UWM-JEPA: Predictive World Models That Imagine in Belief Space

DGX agent

arXiv:2605.25313v1 Announce Type: cross Abstract: World models for partially observed environments must imagine multiple compatible hidden futures and steer between them under counterfactual actions.

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models

DGX agent

arXiv:2602.18788v3 Announce Type: replace Abstract: We introduce BURMESE-SAN, the first holistic benchmark that systematically evaluates large language models (LLMs) for Burmese across three core NLP

model-releasesarxiv-cs-cl
25 May 2026
Applications

Evaluating Customized vs. Generalist Transformer-based Models for Legal Contract Classification

DGX agent

arXiv:2508.07849v2 Announce Type: replace Abstract: Despite advances in legal NLP, no comprehensive evaluation of Transformer-based models customized for legal tasks (referred to as `legal-specific' m

applicationsarxiv-cs-cl
25 May 2026
Model Releases

Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model

DGX agent

arXiv:2605.23790v1 Announce Type: new Abstract: Saliency prediction has been extensively studied in RGB images and videos as a computational model of human visual attention. In contrast, predicting sa

model-releasesarxiv-cs-cv
25 May 2026
Research

InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion

DGX agent

arXiv:2505.13893v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have intensified efforts to fuse heterogeneous open-source models into a unified system that inherit

researcharxiv-cs-cl
25 May 2026
Model Releases

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

DGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

model-releasesarxiv-cs-ai
25 May 2026
Safety

The physics of AI weather models

DGX agent

arXiv:2605.23778v1 Announce Type: cross Abstract: Could it be that AI weather models are solving physical equations, although they may not be the equations used by conventional NWP models? We compute

safetyarxiv-cs-lg
25 May 2026
Model Releases

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models

DGX agent

arXiv:2605.22870v1 Announce Type: cross Abstract: Chain-of-thought (CoT) prompting is necessary for arithmetic in small language models, yet shuffling its steps preserves most performance. What does C

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

The Surprising Difficulty of Search in Model-Based Reinforcement Learning

DGX agent

arXiv:2601.21306v2 Announce Type: replace-cross Abstract: This paper investigates search in model-based reinforcement learning (RL). Conventional wisdom holds that long-term predictions and compoundin

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection

DGX agent

arXiv:2605.21773v1 Announce Type: cross Abstract: Recent benchmark efforts have advanced the evaluation of large language models (LLMs) in cybersecurity, including tasks such as penetration testing an

model-releasesarxiv-cs-lg
23 May 2026
Safety

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

DGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

safetyarxiv-cs-lg
23 May 2026
Model Releases

Truncated Neural Likelihood Estimation for Simulation-Based Inference in State-Space Models

DGX agent

arXiv:2605.21805v1 Announce Type: cross Abstract: State-space models (SSMs) are powerful probabilistic tools for modeling time-varying systems with latent dynamics. Inference in SSMs involves the esti

model-releasesarxiv-cs-lg
23 May 2026
Research

Accelerated Test-Time Scaling with Model-Free Speculative Sampling

DGX agent

arXiv:2506.04708v3 Announce Type: replace Abstract: Language models have demonstrated remarkable capabilities in reasoning tasks through test-time scaling techniques like best-of-N sampling and tree s

researcharxiv-cs-cl
22 May 2026
Model Releases

BEiTScore: Reference-free Image Captioning Evaluation with an Efficient Cross-Encoder Model

DGX agent

arXiv:2605.21728v1 Announce Type: cross Abstract: Image captioning evaluation remains a significant challenge, as vision-language models evolve toward more challenging capabilities such as generating

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models

DGX agent

arXiv:2605.22732v1 Announce Type: cross Abstract: We investigate whether acoustic emotion recognition models can serve as proxies for the Pathos dimension in political speech analysis, as operationali

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

DeepSeek v4: the most expected open-source model ever released, and the quietest landing

DGX agent

After 15 months of incremental updates, leaks, and rumored leaks, DeepSeek released version 4. It arrived without the fanfare R1 and R1-preview commanded in early 2025. That quiet reception is the mos

model-releaseslambda-labs
22 May 2026
Model Releases

GPT 5.5 seems to be improving in that direction now, and Claude models are getting worse at it, so I don't think there's a clear winner now.

DGX agent

Jeremy Howard comments on comparative performance trends between GPT 5.5 and Claude models, noting that GPT 5.5 appears to be improving in a particular capability while Claude models are declining in

model-releasesjeremy-howard--x
22 May 2026
Model Releases

How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing

DGX agent

arXiv:2602.01851v2 Announce Type: replace Abstract: Recent generative models have achieved remarkable progress in image editing. However, existing systems and benchmarks remain largely text-guided. In

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

DGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model

DGX agent

arXiv:2605.22089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework for end-to-end autonomous driving. However, existing VLAs typically rely on sp

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Security Document Classification with a Fine-Tuned Local Large Language Model: Benchmark Data and an Open-Source System

DGX agent

arXiv:2605.20368v1 Announce Type: cross Abstract: Organizations that scan documents for sensitive information face a practical problem. Cloud services require data to be sent to external infrastructur

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Sub-exponential Growth Dynamics in Complex Systems: A Piecewise Power-Law Model for the Diffusion of New Words and Names

DGX agent

arXiv:2511.04106v5 Announce Type: replace-cross Abstract: The diffusion of ideas and language in society has conventionally been described by S-shaped models, such as the logistic curve. However, the

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Teaching Language Models to Forecast Research Success Through Comparative Idea Evaluation

DGX agent

arXiv:2605.21491v1 Announce Type: cross Abstract: As language models accelerate scientific research by automating hypothesis generation and implementation, a new bottleneck emerges: evaluating and fil

model-releasesarxiv-cs-cl
22 May 2026
Research

The Neglected Baseline in Model Interpretation

DGX agent

arXiv:2605.22417v1 Announce Type: new Abstract: We observe that existing model interpretation methods generally ignore the baseline, and such neglect often results in imprecise or even incorrect inter

researcharxiv-cs-cv
22 May 2026
Model Releases

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

DGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Visual-Advantage On-Policy Distillation for Vision-Language Models

DGX agent

arXiv:2605.21924v1 Announce Type: new Abstract: On-policy knowledge distillation has proven effective for language models, yet its application to vision-language models (VLMs) remains underexplored. W

model-releasesarxiv-cs-cv
22 May 2026
Research

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

DGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

researcharxiv-cs-cl
22 May 2026
Model Releases

An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees

DGX agent

arXiv:2605.20521v1 Announce Type: new Abstract: Fine-tuning adapts a pretrained machine learning model to a small, sensitive dataset, but this process risks memorizing individual new data points, maki

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models

DGX agent

arXiv:2605.20777v1 Announce Type: new Abstract: Visual storytelling with diffusion models has made impressive strides in maintaining character consistency across narrative scenes. However, a critical

model-releasesarxiv-cs-cv
21 May 2026
Safety

Bayesian Preference Learning for Test-Time Steerable Reward Models

DGX agent

arXiv:2602.08819v2 Announce Type: replace-cross Abstract: Reward models are central to aligning language models with human preferences via reinforcement learning (RL). As RL is increasingly applied to

safetyarxiv-cs-cl
21 May 2026
Safety

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

DGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

safetyarxiv-cs-cl
21 May 2026
Model Releases

FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation

DGX agent

arXiv:2605.20316v1 Announce Type: new Abstract: Modern text-to-image diffusion models encode rich visual priors, but expose them only through one-way text-conditioned generation. Existing unified visi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation

DGX agent

arXiv:2605.20469v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used for medical image interpretation, yet they frequently hallucinate, generating clinically plausible b

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Optimization Hyper-parameter Laws for Large Language Models

DGX agent

arXiv:2409.04777v4 Announce Type: replace Abstract: Large Language Models have driven significant AI advancements, yet their training is resource-intensive and highly sensitive to hyper-parameter sele

model-releasesarxiv-cs-lg
21 May 2026
Applications

Towards the Anonymization of the Language Modeling

DGX agent

arXiv:2501.02407v3 Announce Type: replace Abstract: Rapid advances in Natural Language Processing (NLP) have revolutionized many fields, including healthcare. However, these advances raise significant

applicationsarxiv-cs-cl
21 May 2026
Model Releases

VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models

DGX agent

arXiv:2605.20774v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong promise for general-purpose robotic manipulation, but their real-world evaluation remains limited

model-releasesarxiv-cs-ro
21 May 2026
← Previous
1…6566676869…1259
Next →