AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Model Releases

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

DGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

DGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

safetyarxiv-cs-lg
12 May 2026
Model Releases

GIFT: Guided Importance-Aware Fine-Tuning for Diffusion Language Models

DGX agent

arXiv:2509.20863v3 Announce Type: replace Abstract: Diffusion models have recently shown strong potential in language modeling, offering faster generation compared to traditional autoregressive approa

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

DGX agent

arXiv:2501.12202v4 Announce Type: replace Abstract: We present Hunyuan3D 2.0, an advanced large-scale 3D synthesis system for generating high-resolution textured 3D assets. This system includes two fo

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Is Your Driving World Model an All-Around Player?

DGX agent

arXiv:2605.10858v1 Announce Type: new Abstract: Today's driving world models can generate remarkably realistic dash-cam videos, yet no single model excels universally. Some generate photorealistic tex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Model-Free Neural Filtering: A Comparison with Classical Filters in Nonlinear Systems

DGX agent

arXiv:2601.21266v3 Announce Type: replace Abstract: Neural network models are increasingly used for state estimation in control and decision-making, yet it remains unclear to what extent they behave a

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Benchmarking World-Model Learning with Environment-Level Queries

DGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Fine-tuning a vision-language model for fracture-surface morphology recognition

DGX agent

arXiv:2605.07145v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong potential for scientific image understanding, but general-purpose models often lack the domain-specifi

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Domain Incremental Continual Learning Benchmark for ICU Time Series Model Transportability

DGX agent

arXiv:2605.03832v1 Announce Type: new Abstract: In recent years, machine learning has made significant progress in clinical outcome prediction, demonstrating increasingly accurate results. However, th

model-releasesarxiv-cs-lg
6 May 2026
Research

Mechanism-Faithful Queueing Simulation Model Translation with Large Language Model Support

DGX agent

arXiv:2601.06543v2 Announce Type: replace Abstract: Queueing simulation studies often require substantial manual effort to translate conceptual system descriptions into executable programs and to veri

researcharxiv-cs-cl
6 May 2026
Model Releases

StateVLM: A State-Aware Vision-Language Model for Robotic Affordance Reasoning

DGX agent

arXiv:2605.03927v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown remarkable performance in various robotic tasks, as they can perceive visual information and understand natural

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

CNN-based Multi-In-Multi-Out Model for Efficient Spatiotemporal Prediction

DGX agent

arXiv:2605.01277v1 Announce Type: new Abstract: Recently, Convolutional Neural Network (CNN) or Transformer architecture based models have been proposed to overcome the limitations of Recurrent Neural

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Dispersion Loss Counteracts Embedding Condensation and Improves Generalization in Small Language Models

DGX agent

arXiv:2602.00217v2 Announce Type: replace Abstract: Large language models (LLMs) achieve remarkable performance through ever-increasing parameter counts, but scaling incurs steep computational costs.

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Fine-Tuning Impairs the Balancedness of Foundation Models in Long-tailed Personalized Federated Learning

DGX agent

arXiv:2605.02247v1 Announce Type: new Abstract: Personalized federated learning (PFL) with foundation models has emerged as a promising paradigm enabling clients to adapt to heterogeneous data distrib

model-releasesarxiv-cs-cv
5 May 2026
Tutorials

GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models

DGX agent

arXiv:2605.01256v1 Announce Type: new Abstract: A promising paradigm for adapting instruction-tuned language models is to learn task-specific updates on a pretrained base model and subsequently merge

tutorialsarxiv-cs-cl
5 May 2026
Research

The Pre-Training Study of Expanded-SPLADE Models on Web Document Titles

DGX agent

arXiv:2605.01407v1 Announce Type: cross Abstract: Masked Language Modeling (MLM) pre-training is one of the primary ways to initialize Neural Information Retrieval (IR) models prior to retrieval fine-

researcharxiv-cs-cl
5 May 2026
Model Releases

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

DGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks

DGX agent

arXiv:2507.01955v3 Announce Type: replace Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed visual understanding beyond que

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Jailbroken Frontier Models Retain Their Capabilities

DGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

model-releasesarxiv-cs-lg
4 May 2026
Tutorials

haha our model likes to talk about goblins no of course we dont know why, we dont know why the model does anything - yes we are trying to make a superintelligent machine god, maybe it will like goblins too, we have no way of knowing what it will like, we hope it will like humans

DGX agent

This Reddit post from r/ChatGPT humorously discusses an AI model's unexplained tendency to frequently mention goblins in its outputs, using this quirk to reflect on broader uncertainties around AI beh

tutorialsr-chatgpt
1 May 2026
Safety

MotuBrain: An Advanced World Action Model for Robot Control

DGX agent

arXiv:2604.27792v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models achieve strong semantic generalization but often lack fine-grained modeling of world dynamics. Recent work explores

safetyarxiv-cs-ro
1 May 2026
Research

Learning Illumination Control in Diffusion Models

DGX agent

arXiv:2604.24877v1 Announce Type: new Abstract: Controlling illumination in images is essential for photography and visual content creation. While closed-source models have demonstrated impressive ill

researcharxiv-cs-cv
29 Apr 2026
Research

Revisiting the Past: Data Unlearning with Model State History

DGX agent

arXiv:2506.20941v3 Announce Type: replace Abstract: Large language models are trained on massive corpora of web data, which may include private data, copyrighted material, factually inaccurate data, o

researcharxiv-cs-lg
29 Apr 2026
Model Releases

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving

DGX agent

arXiv:2604.22851v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have advanced highlevel reasoning in autonomous driving, their ability to ground this reasoning in the underlying

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Evaluating whether AI models would sabotage AI safety research

DGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation

DGX agent

arXiv:2604.00829v3 Announce Type: replace-cross Abstract: Adapting pretrained language models (LMs) into vision-language models (VLMs) can degrade their native linguistic capability due to representat

safetyarxiv-cs-cl
28 Apr 2026
Research

On the Memorization of Consistency Distillation for Diffusion Models

DGX agent

arXiv:2604.23552v1 Announce Type: cross Abstract: Diffusion models are central to modern generative modeling, and understanding how they balance memorization and generalization is critical for reliabl

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where y…

DGX agent

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where you don't need frontier intelligence. you want cheap, fast, an

model-releasesharrison-chase--x
28 Apr 2026
Model Releases

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

DGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?

DGX agent

arXiv:2510.10254v2 Announce Type: replace Abstract: Recent advances in large generative models have shown that simple autoregressive formulations, when scaled appropriately, can exhibit strong zero-sh

researcharxiv-cs-cv
27 Apr 2026
Model Releases

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden…

DGX agent

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden to some lab publishing weights. but i take that over compan

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

DeepSeek V4 - almost on the frontier, a fraction of the price

DGX agent

Chinese AI lab DeepSeek's last model release was V3.2 (and V3.2 Speciale) last December. They just dropped the first of their hotly anticipated V4 series in the shape of two preview models, DeepSeek-V

model-releasessimon-willison
24 Apr 2026
Model Releases

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

DGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Model page for more information and integrations: https://ollama.com/library/deepseek-v4-flash

DGX agent

Ollama announced DeepSeek-v4-flash, a lightweight variant of the DeepSeek-v4 model, now available in their model library for local deployment and integration. The model page provides documentation, us

model-releasesollama--x
24 Apr 2026
Model Releases

OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model

DGX agent

arXiv:2604.20806v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have made substantial advances in reasoning tasks at the Olympiad level. Nevertheless, current Olympiad-level mul

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Surrogate modeling for interpreting black-box LLMs in medical predictions

DGX agent

arXiv:2604.20331v1 Announce Type: cross Abstract: Large language models (LLMs), trained on vast datasets, encode extensive real-world knowledge within their parameters, yet their black-box nature obsc

applicationsarxiv-cs-ai
23 Apr 2026
Applications

Beyond Coefficients: Forecast-Necessity Testing for Interpretable Causal Discovery in Nonlinear Time-Series Models

DGX agent

arXiv:2604.18751v1 Announce Type: cross Abstract: Nonlinear machine-learning models are increasingly used to discover causal relationships in time-series data, yet the interpretation of their outputs

applicationsarxiv-cs-ai
22 Apr 2026
Applications

Handling and Interpreting Missing Modalities in Patient Clinical Trajectories via Autoregressive Sequence Modeling

DGX agent

arXiv:2604.18753v1 Announce Type: cross Abstract: An active challenge in developing multimodal machine learning (ML) models for healthcare is handling missing modalities during training and deployment

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

IndiaFinBench: An Evaluation Benchmark for Large Language Model Performance on Indian Financial Regulatory Text

DGX agent

arXiv:2604.19298v1 Announce Type: cross Abstract: We introduce IndiaFinBench, to our knowledge the first publicly available evaluation benchmark for assessing large language model (LLM) performance on

model-releasesarxiv-cs-ai
22 Apr 2026
Research

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

DGX agent

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffus

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Finding Culture-Sensitive Neurons in Vision-Language Models

DGX agent

arXiv:2510.24942v2 Announce Type: replace-cross Abstract: Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs proce

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models

DGX agent

arXiv:2604.17941v1 Announce Type: cross Abstract: Recent work has increasingly explored neuron-level interpretation in vision-language models (VLMs) to identify neurons critical to final predictions.

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?

DGX agent

arXiv:2604.17930v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit a puzzling disparity in their formal linguistic competence: while they learn some linguistic phenomena with near-pe

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

DGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Reciprocal Co-Training (RCT): Coupling Gradient-Based and Non-Differentiable Models via Reinforcement Learning

DGX agent

arXiv:2604.16378v1 Announce Type: new Abstract: Large language models (LLMs) and classical machine learning methods offer complementary strengths for predictive modeling, yet their fundamentally diffe

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

DGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection

DGX agent

arXiv:2508.11281v3 Announce Type: replace Abstract: Detecting toxic content using language models is crucial yet challenging. While substantial progress has been made in English, toxicity detection in

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Understanding Counting Mechanisms in Large Language and Vision-Language Models

DGX agent

arXiv:2511.17699v2 Announce Type: replace Abstract: Counting is one of the fundamental abilities of large language models (LLMs) and large vision-language models (LVLMs). This paper examines how these

researcharxiv-cs-cv
21 Apr 2026
← Previous
1…2324252627…1247
Next →