AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models

DGX agent

arXiv:2605.19137v1 Announce Type: new Abstract: Video foundation models achieve strong performance across many video understanding tasks, but typically require large-scale pre-training on massive vide

researcharxiv-cs-cv
20 May 2026
Tutorials
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

WIND: Weather Inverse Diffusion for Zero-Shot Atmospheric Modeling

DGX agent

arXiv:2602.03924v2 Announce Type: replace-cross Abstract: Deep learning has revolutionized weather forecasting, but many challenges remain, including climate modeling. Moreover, the current landscape

tutorialsarxiv-cs-ai
20 May 2026
Model Releases

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback

DGX agent

arXiv:2605.18073v1 Announce Type: cross Abstract: Large Language Models (LLMs) demonstrate strong potential for automated code generation, yet their ability to iteratively refine solutions using execu

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

AgroCoT: A Chain-of-Thought Benchmark for Evaluating Reasoning in Vision-Language Models for Agriculture

DGX agent

arXiv:2511.23253v3 Announce Type: replace Abstract: Recent advancements in Vision-Language Models (VLMs) have significantly impacted various industries. In agriculture, these multimodal capabilities h

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Ancient Greek to Modern Greek Machine Translation: A Novel Benchmark and Fine-Tuning Experiments on LLMs and NMT Models

DGX agent

arXiv:2605.18504v1 Announce Type: new Abstract: Machine Translation (MT) for Ancient Greek (AG) to Modern Greek (MG) is a low-resource task, constrained by the lack of large-scale, high-quality parall

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Beacon: Single-Turn Diagnosis and Mitigation of Latent Sycophancy in Large Language Models

DGX agent

arXiv:2510.16727v2 Announce Type: replace-cross Abstract: Large language models internalize a structural trade-off between truthfulness and obsequious flattery, emerging from reward optimization that

model-releasesarxiv-cs-ai
19 May 2026
Safety

CatalyticMLLM: A Graph-Text Multimodal Large Language Model for Catalytic Materials

DGX agent

arXiv:2605.17254v1 Announce Type: new Abstract: Property prediction and inverse structural design of catalytic materials are typically modeled as two independent tasks: the former predicts target prop

safetyarxiv-cs-ai
19 May 2026
Safety

Factored Causal Representation Learning for Robust Reward Modeling in RLHF

DGX agent

arXiv:2601.21350v2 Announce Type: replace Abstract: A reliable reward model is essential for aligning large language models with human preferences through reinforcement learning from human feedback. H

safetyarxiv-cs-lg
19 May 2026
Model Releases

Fine-tuning Pocket-Aware Diffusion Models via Denoising Policy Optimization

DGX agent

arXiv:2605.17693v1 Announce Type: cross Abstract: Structure-based drug design has been accelerated by pocket-aware 3D generative models, yet most methods primarily fit the training distribution and ma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

FormuLLA: A Large Language Model Approach to Generating Novel 3D Printable Formulations

DGX agent

arXiv:2601.02071v3 Announce Type: replace Abstract: Pharmaceutical three-dimensional (3D) printing is an advanced fabrication technology with the potential to enable truly personalised dosage forms. R

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets

DGX agent

arXiv:2605.18475v1 Announce Type: cross Abstract: Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. Ho

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

HEED: Density-Weighted Residual Alignment for Hybrid Vision-Language Model Distillation

DGX agent

arXiv:2605.17093v1 Announce Type: cross Abstract: Distilling vision-language models into faster hybrid architectures, such as 3:1 Mamba-2/attention mixes, is now standard practice for making inference

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Hunt Instead of Wait: Evaluating Deep Data Research on Large Language Models

DGX agent

arXiv:2602.02039v2 Announce Type: replace Abstract: The agency expected of Agentic Large Language Models goes beyond answering correctly, requiring autonomy to set goals and decide what to explore. We

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Identifiable Token Correspondence for World Models

DGX agent

arXiv:2605.16457v1 Announce Type: cross Abstract: Transformer-based world models have shown strong performance in visual reinforcement learning, but often suffer from temporal inconsistency in long-ho

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

KairosHope: A Next-Generation Time-Series Foundation Model for Specialized Classification via Dual-Memory Architecture

DGX agent

arXiv:2605.18657v1 Announce Type: cross Abstract: Time Series Foundation Models (TSFMs) have demonstrated notable success in general-purpose forecasting tasks; however, their adaptation to specialized

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

LAST-RAG: Literature-Anchored Stochastic Trajectory Retrieval-Augmented Generation for Knowledge-Conditioned Degradation Model Selection

DGX agent

arXiv:2605.17902v1 Announce Type: new Abstract: Stochastic-process-based degradation modeling is a core approach for estimating the distribution of remaining useful life (RUL); however, the selection

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting

DGX agent

arXiv:2510.10528v3 Announce Type: replace Abstract: Large reasoning models (LRMs) have demonstrated remarkable proficiency in tackling complex tasks through step-by-step thinking. However, this length

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models

DGX agent

arXiv:2605.16409v1 Announce Type: cross Abstract: Optical character recognition (OCR) and multilingual text understanding remain major failure modes of multimodal large language models (MLLMs), partic

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Seeing Together:Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models

DGX agent

arXiv:2605.18431v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made substantial progress in egocentric video understanding, but their ability to reason cooperatively fro

model-releasesarxiv-cs-cv
19 May 2026
Research

Statistical Hand Shape Modeling from Clinical CT Scans Using Deep Learning and Implicit Skinning

DGX agent

arXiv:2605.16980v1 Announce Type: new Abstract: Accurate segmentation and statistical shape modeling of hand anatomy have significant implications for medical diagnostics, ergonomics, and biomechanics

researcharxiv-cs-cv
19 May 2026
Model Releases

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning

DGX agent

arXiv:2602.10503v2 Announce Type: replace Abstract: Pretrained on large-scale and diverse datasets, VLA models demonstrate strong generalization and adaptability as general-purpose robotic policies. H

model-releasesarxiv-cs-ro
19 May 2026
Research

VideoNeuMat: Neural Material Extraction from Generative Video Models

DGX agent

arXiv:2602.07272v2 Announce Type: replace Abstract: Creating photorealistic materials for 3D rendering requires exceptional artistic skill. Generative models for materials could help, but are currentl

researcharxiv-cs-cv
19 May 2026
Research

Why Do Reasoning Models Lose Coverage? The Role of Data and Forks in the Road

DGX agent

arXiv:2605.17026v1 Announce Type: new Abstract: Recent progress in large language models has led to the emergence of reasoning models, which have shown strong performance on complex tasks through spec

researcharxiv-cs-lg
19 May 2026
Model Releases

Evaluating Chinese Ambiguity Understanding in Large Language Models

DGX agent

arXiv:2605.15635v1 Announce Type: new Abstract: Linguistic ambiguity is critical to the robustness of Large Language Models (LLMs), yet existing research focuses mostly on English, with limited attent

model-releasesarxiv-cs-cl
18 May 2026
Safety

f-Trajectory Balance: A Loss Family for Tuning GFlowNets, Generative Models, and LLMs with Off- and On-Policy Data

DGX agent

arXiv:2605.15417v1 Announce Type: cross Abstract: In GFlowNets and variational inference, it has been shown that the mean square error between target and model log probabilities is an effective, low v

safetyarxiv-cs-ai
18 May 2026
Model Releases

FINESSE-Bench: A Hierarchical Benchmark Suite for Financial Domain Knowledge and Technical Analysis in Large Language Models

DGX agent

arXiv:2605.15482v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being applied to financial analysis, reporting, investment decision support, risk management, compliance,

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Frontier Large Language Models Rival State-of-the-Art Planners

DGX agent

arXiv:2511.09378v2 Announce Type: replace Abstract: A series of influential studies established that large language models cannot reliably solve even simple planning tasks. We show that the latest gen

model-releasesarxiv-cs-ai
18 May 2026
Research

Learning Normalized Energy Models for Linear Inverse Problems

DGX agent

arXiv:2605.15487v1 Announce Type: cross Abstract: Generative diffusion models can provide powerful prior probability models for inverse problems in imaging, but existing implementations suffer from tw

researcharxiv-cs-cv
18 May 2026
Model Releases

MHGraphBench: Knowledge Graph-Grounded Benchmarking of Mental Health Knowledge in Large Language Models

DGX agent

arXiv:2605.15589v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in the mental health domain, yet it remains unclear how well they capture related biomedical knowledg

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Social-Mamba: Socially-Aware Trajectory Forecasting with State-Space Models

DGX agent

arXiv:2605.15424v1 Announce Type: new Abstract: Human trajectory forecasting is crucial for safe navigation in crowded environments, requiring models that balance accuracy with computational efficienc

model-releasesarxiv-cs-cv
18 May 2026
Research

Time-Varying Deep State Space Models for Sequences with Switching Dynamics

DGX agent

arXiv:2605.15311v1 Announce Type: new Abstract: The identification and modeling of time-varying systems is a fundamental challenge in signal processing and system identification. To address this chall

researcharxiv-cs-lg
18 May 2026
Model Releases

Zero-Shot Goal Recognition with Large Language Models

DGX agent

arXiv:2605.15333v1 Announce Type: new Abstract: Large language models have recently reached near-parity with classical planners on well-known planning domains, yet this competence relies on world-know

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Eradicating Negative Transfer in Multi-Physics Foundation Models via Sparse Mixture-of-Experts Routing

DGX agent

arXiv:2605.15179v1 Announce Type: cross Abstract: Scaling Scientific Machine Learning (SciML) toward universal foundation models is bottlenecked by negative transfer: the simultaneous co-training of d

model-releasesarxiv-cs-ai
15 May 2026
Applications

Finding Interpretable Prompt-Specific Circuits in Language Models

DGX agent

arXiv:2602.13483v2 Announce Type: replace-cross Abstract: Understanding the internal circuits that language models use to solve tasks remains a central challenge in mechanistic interpretability. A cru

applicationsarxiv-cs-ai
15 May 2026
Model Releases

MechVerse: Evaluating Physical Motion Consistency in Video Generation Models

DGX agent

arXiv:2605.14843v1 Announce Type: new Abstract: Text- and image-conditioned video generation models have achieved strong visual fidelity and temporal coherence, but they often fail to generate motion

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

VectraYX-Nano: A 42M-Parameter Spanish Cybersecurity Language Model with Curriculum Learning and Native Tool Use

DGX agent

arXiv:2605.13989v1 Announce Type: new Abstract: We present VectraYX-Nano, a 41.95M-parameter decoder-only language model trained from scratch in Spanish for cybersecurity, with a Latin-American focus

model-releasesarxiv-cs-cl
15 May 2026
Agents

AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents

DGX agent

arXiv:2605.13357v1 Announce Type: cross Abstract: Foundation models have transformed automated code generation, yet autonomous software-engineering agents remain unreliable in realistic development se

agentsarxiv-cs-ai
14 May 2026
Research

Asymmetric Flow Models

DGX agent

arXiv:2605.12964v1 Announce Type: new Abstract: Flow-based generation in high-dimensional spaces is difficult because velocity prediction requires modeling high-dimensional noise, even when data has s

researcharxiv-cs-cv
14 May 2026
Model Releases

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

DGX agent

arXiv:2605.13773v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether t

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

How Well Do Large-Scale Chemical Language Models Transfer to Downstream Tasks?

DGX agent

arXiv:2602.11618v4 Announce Type: replace Abstract: Chemical Language Models (CLMs) pre-trained on large scale molecular data are widely used for molecular property prediction. However, the common bel

model-releasesarxiv-cs-lg
14 May 2026
Safety

Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling

DGX agent

arXiv:2605.13801v1 Announce Type: cross Abstract: As generative AI models such as large language models (LLMs) become more pervasive, ensuring the safety, robustness, and overall trustworthiness of th

safetyarxiv-cs-ai
14 May 2026
Model Releases

Inducing Overthink: Hierarchical Genetic Algorithm-based DoS Attack on Black-Box Large Language Reasoning Models

DGX agent

arXiv:2605.13338v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) are increasingly integrated into systems requiring reliable multi-step inference, yet this growing dependence exposes ne

model-releasesarxiv-cs-ai
14 May 2026
Safety

Integration of an Agent Model into an Open Simulation Architecture for Scenario-Based Testing of Automated Vehicles

DGX agent

arXiv:2605.13539v1 Announce Type: new Abstract: Simulative and scenario-based testing are crucial methods in the safety assurance for automated driving systems. To ensure that simulation results are r

safetyarxiv-cs-ro
14 May 2026
Model Releases

Large Language Models Lack Temporal Awareness of Medical Knowledge

DGX agent

arXiv:2605.13045v1 Announce Type: new Abstract: The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, w

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models

DGX agent

arXiv:2505.15616v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved significant advances in integrating visual and linguistic information, yet their ability to r

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Probing Persona-Dependent Preferences in Language Models

DGX agent

arXiv:2605.13339v1 Announce Type: cross Abstract: Large language models (LLMs) can be said to have preferences: they reliably pick certain tasks and outputs over others, and preferences shaped by post

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SupChain-Bench: Benchmarking Large Language Models for Real-World Supply Chain Management

DGX agent

arXiv:2602.07342v2 Announce Type: replace Abstract: Large language models (LLMs) have shown promise in complex reasoning and tool-based decision making, motivating their application to real-world supp

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

The critical slowing down in diffusion models

DGX agent

arXiv:2605.12597v1 Announce Type: cross Abstract: Computational sampling has been central to the sciences since the mid-20th century. While machine-learning-based approaches have recently enabled majo

model-releasesarxiv-cs-ai
14 May 2026
← Previous
1…6768697071…1030
Next →