(1)
An In-Depth Comparison of Plain CNN, Fine-Tune VGG16, and Vision Transformer Models in Object Detection. BJET 2026, 21 (1), 46-54. https://doi.org/10.67607/mwyr5z74.