<delect id="aqmse"><source id="aqmse"></source></delect>

<center id="aqmse"><strike id="aqmse"></strike></center>

<tr id="aqmse"><blockquote id="aqmse"></blockquote></tr><rt id="aqmse"><abbr id="aqmse"></abbr></rt>^{<bdo id="aqmse"></bdo>}

溫馨提示×

溫馨提示×

您好，登錄后才能下訂單哦！

密碼登錄×

忘記密碼？

登錄注冊×

獲取短信驗證碼

其他方式登錄

點擊登錄注冊即表示同意《億速云用戶服務條款》

用戶登錄×

賬戶密碼登錄

請使用微信掃描上方二維碼

使用幫助

請求超時！

請點擊重新獲取二維碼

TensorFlow中如何讀取圖像數(shù)據(jù)

發(fā)布時間：2020-07-01 10:56:57 來源：億速云閱讀：149 作者：清晨欄目：開發(fā)技術(shù)

小編給大家分享一下TensorFlow中如何讀取圖像數(shù)據(jù)，希望大家閱讀完這篇文章后大所收獲，下面讓我們一起去探討方法吧！

　三種讀取數(shù)據(jù)的方式，分別用于處理單張圖片、大量圖片，和TFRecorder讀取方式。并且還補充了功能相近的tf函數(shù)。

1、處理單張圖片

　　我們訓練完模型之后，常常要用圖片測試，有的時候，我們并不需要對很多圖像做測試，可能就是幾張甚至一張。這種情況下沒有必要用隊列機制。

import tensorflow as tf
import matplotlib.pyplot as plt

def read_image(file_name):
 img = tf.read_file(filename=file_name)  # 默認讀取格式為uint8
 print("img 的類型是",type(img));
 img = tf.image.decode_jpeg(img,channels=0) # channels 為1得到的是灰度圖，為0則按照圖片格式來讀
 return img

def main( ):
 with tf.device("/cpu:0"):
　　　　  # img_path是文件所在地址包括文件名稱，地址用相對地址或者絕對地址都行 
   img_path='./1.jpg'
   img=read_image(img_path)
   with tf.Session() as sess:
   image_numpy=sess.run(img)
   print(image_numpy)
   print(image_numpy.dtype)
   print(image_numpy.shape)
   plt.imshow(image_numpy)
   plt.show()

if __name__=="__main__":
 main()

"""

輸出結(jié)果為：

img 的類型是 <class 'tensorflow.python.framework.ops.Tensor'>
[[[196 219 209]
[196 219 209]
[196 219 209]
...
[[ 71 106 42]
[ 59 89 39]
[ 34 63 19]
...
[ 21 52 46]
[ 15 45 43]
[ 22 50 53]]]
uint8
(675, 1200, 3)
"""

　　和tf.read_file用法相似的函數(shù)還有tf.gfile.FastGFile tf.gfile.GFile，只是要指定讀取方式是'r' 還是'rb' 。

2、需要讀取大量圖像用于訓練

　　這種情況就需要使用Tensorflow隊列機制。首先是獲得每張圖片的路徑，把他們都放進一個list里面，然后用string_input_producer創(chuàng)建隊列，再用tf.WholeFileReader讀取。具體請看下例：

def get_image_batch(data_file,batch_size):
 data_names=[os.path.join(data_file,k) for k in os.listdir(data_file)]
 
 #這個num_epochs函數(shù)在整個Graph是local Variable，所以在sess.run全局變量的時候也要加上局部變量。 
 filenames_queue=tf.train.string_input_producer(data_names,num_epochs=50,shuffle=True,capacity=512)
 reader=tf.WholeFileReader()
 _,img_bytes=reader.read(filenames_queue)
 image=tf.image.decode_png(img_bytes,channels=1) #讀取的是什么格式，就decode什么格式
 #解碼成單通道的，并且獲得的結(jié)果的shape是[&#63;, &#63;,1]，也就是Graph不知道圖像的大小，需要set_shape
 image.set_shape([180,180,1]) #set到原本已知圖像的大小?；蛘咧苯油ㄟ^tf.image.resize_images，tf.reshape()
 image=tf.image.convert_image_dtype(image,tf.float32)
 #預處理 下面的一句代碼可以換成自己想使用的預處理方式
 #image=tf.divide(image,255.0) 
 return tf.train.batch([image],batch_size)

　　這里的date_file是指文件夾所在的路徑，不包括文件名。第一句是遍歷指定目錄下的文件名稱，存放到一個list中。當然這個做法有很多種方法，比如glob.glob，或者tf.train.match_filename_once

全部代碼如下：

import tensorflow as tf
import os
def read_image(data_file,batch_size):
 data_names=[os.path.join(data_file,k) for k in os.listdir(data_file)]
 filenames_queue=tf.train.string_input_producer(data_names,num_epochs=5,shuffle=True,capacity=30)
 reader=tf.WholeFileReader()
 _,img_bytes=reader.read(filenames_queue)
 image=tf.image.decode_jpeg(img_bytes,channels=1)
 image=tf.image.resize_images(image,(180,180))

 image=tf.image.convert_image_dtype(image,tf.float32)
 return tf.train.batch([image],batch_size)

def main( ):
 img_path=r'F:\dataSet\WIDER\WIDER_train\images\6--Funeral' #本地的一個數(shù)據(jù)集目錄，有足夠的圖像
 img=read_image(img_path,batch_size=10)
 image=img[0] #取出每個batch的第一個數(shù)據(jù)
 print(image)
 init=[tf.global_variables_initializer(),tf.local_variables_initializer()]
 with tf.Session() as sess:
  sess.run(init)
  coord = tf.train.Coordinator()
  threads = tf.train.start_queue_runners(sess=sess,coord=coord)
  try:
   while not coord.should_stop():
    print(image.shape)
  except tf.errors.OutOfRangeError:
   print('read done')
  finally:
   coord.request_stop()
  coord.join(threads)


if __name__=="__main__":
 main()

"""

輸出如下：

(180, 180, 1)
(180, 180, 1)
(180, 180, 1)
(180, 180, 1)
(180, 180, 1)
"""

　　這段代碼可以說寫的很是規(guī)整了。注意到init里面有對local變量的初始化，并且因為用到了隊列，當然要告訴電腦什么時候隊列開始, tf.train.Coordinator 和 tf.train.start_queue_runners 就是兩個管理隊列的類，用法如程序所示。

　　與 tf.train.string_input_producer相似的函數(shù)是 tf.train.slice_input_producer。 tf.train.slice_input_producer和tf.train.string_input_producer的第一個參數(shù)形式不一樣。等有時間再做一個二者比較的博客

3、對TFRecorder解碼獲得圖像數(shù)據(jù)

　　其實這塊和上一種方式差不多的，更重要的是怎么生成TFRecorder文件，這一部分我會補充到另一篇博客上。

　　仍然使用 tf.train.string_input_producer。

import tensorflow as tf
import matplotlib.pyplot as plt
import os
import cv2
import numpy as np
import glob

def read_image(data_file,batch_size):
 files_path=glob.glob(data_file)
 queue=tf.train.string_input_producer(files_path,num_epochs=None)
 reader = tf.TFRecordReader()
 print(queue)
 _, serialized_example = reader.read(queue)
 features = tf.parse_single_example(
  serialized_example,
  features={
   'image_raw': tf.FixedLenFeature([], tf.string),
   'label_raw': tf.FixedLenFeature([], tf.string),
  })
 image = tf.decode_raw(features['image_raw'], tf.uint8)
 image = tf.cast(image, tf.float32)
 image.set_shape((12*12*3))
 label = tf.decode_raw(features['label_raw'], tf.float32)
 label.set_shape((2))
 # 預處理部分省略，大家可以自己根據(jù)需要添加
 return tf.train.batch([image,label],batch_size=batch_size,num_threads=4,capacity=5*batch_size)

def main( ):
 img_path=r'F:\python\MTCNN_by_myself\prepare_data\pnet*.tfrecords' #本地的幾個tf文件
 img,label=read_image(img_path,batch_size=10)
 image=img[0]
 init=[tf.global_variables_initializer(),tf.local_variables_initializer()]
 with tf.Session() as sess:
  sess.run(init)
  coord = tf.train.Coordinator()
  threads = tf.train.start_queue_runners(sess=sess,coord=coord)
  try:
   while not coord.should_stop():
    print(image.shape)
  except tf.errors.OutOfRangeError:
   print('read done')
  finally:
   coord.request_stop()
  coord.join(threads)


if __name__=="__main__":
 main()

　　在read_image函數(shù)中，先使用glob函數(shù)獲得了存放tfrecord文件的列表，然后根據(jù)TFRecord文件是如何存的就如何parse，再set_shape；這里有必要提醒下parse的方式。我們看到這里用的是tf.decode_raw ，因為做TFRecord是將圖像數(shù)據(jù)string化了，數(shù)據(jù)是串行的，丟失了空間結(jié)果。從features中取出image和label的數(shù)據(jù)，這時就要用 tf.decode_raw 解碼，得到的結(jié)果當然也是串行的了，所以set_shape 成一個串行的，再reshape。這種方式是取決于你的編碼TFRecord方式的。

再舉一種例子：

reader=tf.TFRecordReader()
_,serialized_example=reader.read(file_name_queue)
features = tf.parse_single_example(serialized_example, features={
 'data': tf.FixedLenFeature([256,256], tf.float32), ###
 'label': tf.FixedLenFeature([], tf.int64),
 'id': tf.FixedLenFeature([], tf.int64)
})
img = features['data']
label =features['label']
id = features['id']

　　這個時候就不需要任何解碼了。因為做TFRecord的方式就是直接把圖像數(shù)據(jù)append進去了。

看完了這篇文章，相信你對TensorFlow中如何讀取圖像數(shù)據(jù)有了一定的了解，想了解更多相關(guān)知識，歡迎關(guān)注億速云行業(yè)資訊頻道，感謝各位的閱讀！

向AI問一下細節(jié)

推薦閱讀：

免責聲明：本站發(fā)布的內(nèi)容（圖片、視頻和文字）以原創(chuàng)、轉(zhuǎn)載和分享為主，文章觀點不代表本網(wǎng)站立場，如果涉及侵權(quán)請聯(lián)系站長郵箱：is@yisu.com進行舉報，并提供相關(guān)證據(jù)，一經(jīng)查實，將立刻刪除涉嫌侵權(quán)內(nèi)容。

上一篇新聞：
SQL Server數(shù)據(jù)庫鏡像搭建(無見證無域控)
下一篇新聞：
php中的this關(guān)鍵字有什么用

猜你喜歡

AI
助
手

產(chǎn)品服務

地區(qū)劃分

專題活動

幫助支持

關(guān)于我們

售后咨詢

7*24小時在線電話：400-100-2938

7*24小時在線 QQ：800811969

關(guān)注億速云

億速云公眾號

手機網(wǎng)站二維碼

<tbody id="uiksg"><sup id="uiksg"></sup></tbody>

<center id="uiksg"><strike id="uiksg"></strike></center>

<object id="uiksg"><code id="uiksg"></code></object>