Repository logo
Log In(current)
  1. Home
  2. Colleges & Schools
  3. Graduate School
  4. Doctoral Dissertations
  5. Object Detection Models with Limited or Small Labeled Datasets
Details

Object Detection Models with Limited or Small Labeled Datasets

Date Issued
December 1, 2023
Author(s)
Fallas Moya, Fabian Enrique  
Advisor(s)
Hairong Qi
Additional Advisor(s)
Hairong Qi
Amir Sadovnik
Hao Gan
Sai Swaminathan
Permanent URI
https://trace.tennessee.edu/handle/20.500.14382/30277
Abstract

Deep Learning made a substantial improvement in the results of different computer vision tasks. However, the use of deep learning models comes with a significant cost: the requirement for substantial amounts of labeled data, also known as annotated data.


Different techniques have been developed in order to train deep learning models with a limited number of labeled data samples. In image classification, this problem - of limited labeled data - has been addressed with great success, but in object detection, the problem is more complex given its nature where two learning tasks are performed: the localization of the boxes - a regression step - and the classification of each box, which makes object detection more complicated.

The main concern in such models revolves around the prudent utilization of labeled samples to yield optimal results when performing inferences. Addressing this issue is a hard challenge and is the core of this dissertation, which delves into diverse strategies for harnessing labeled data to enhance model performance.

We investigated three methods to enhance performance when limited labeled data is available. First, we looked at how to preprocess the available data using multispectral channels, metadata, and data augmentation to increase performance. Secondly, we studied semi-supervised learning in object detection by refining the pseudo-labels using additional models before and after the non-maximum suppression (NMS) method. Lastly, we explored the concept of direct inference object detection, using a foundational visual model (VFM) and few-shot models (from the labeled dataset) to make direct inferences without any training steps.

In all cases, we show how there is a significant improvement, shedding light on potential directions for future research in this domain.

Subjects

object detection

semi-supervised learn...

preprocessing

few-shot learning

foundation model

Disciplines
Computational Engineering
Degree
Doctor of Philosophy
Major
Computer Science
Embargo Date
December 15, 2025
File(s)
Thumbnail Image
Name

Dissertation_defense_Fabian__3_.pdf

Size

17.57 MB

Format

Adobe PDF

Checksum (MD5)

2f1c21c2c1920303560f40a0051cede3


University Libraries

1015 Volunteer Boulevard
Knoxville, TN 37996
865-974-4351

Map & Directions
Donate to the Libraries
  • About
  • John C. Hodges Society
  • Speaking Volumes magazine
  • Outreach
  • Directory
  • Employment
  • Policies
  • Library Intranet
University of Tennessee power T logo

The University of Tennessee, Knoxville
Knoxville, Tennessee 37996
865-974-1000

Events
A-Z
Apply
Privacy
Map
Directory
Give to UT
Accessibility

Built with DSpace-CRIS software - Extension maintained and optimized by 4Science