關於tensorflow softmax函式用法解析

阿新 • • 發佈：2020-07-01

如下所示：

def softmax(logits,axis=None,name=None,dim=None):
 """Computes softmax activations.
 This function performs the equivalent of
  softmax = tf.exp(logits) / tf.reduce_sum(tf.exp(logits),axis)
 Args:
 logits: A non-empty `Tensor`. Must be one of the following types: `half`,`float32`,`float64`.
 axis: The dimension softmax would be performed on. The default is -1 which
  indicates the last dimension.
 name: A name for the operation (optional).
 dim: Deprecated alias for `axis`.
 Returns:
 A `Tensor`. Has the same type and shape as `logits`.
 Raises:
 InvalidArgumentError: if `logits` is empty or `axis` is beyond the last
  dimension of `logits`.
 """
 axis = deprecation.deprecated_argument_lookup("axis",axis,"dim",dim)
 if axis is None:
 axis = -1
 return _softmax(logits,gen_nn_ops.softmax,name)

softmax函式的返回結果和輸入的tensor有相同的shape，既然沒有改變tensor的形狀，那麼softmax究竟對tensor做了什麼？

答案就是softmax會以某一個軸的下標為索引，對這一軸上其他維度的值進行啟用 + 歸一化處理。

一般來說，這個索引軸都是表示類別的那個維度（tf.nn.softmax中預設為axis=-1,也就是最後一個維度）

舉例：

def softmax(X,theta = 1.0,axis = None):
 """
 Compute the softmax of each element along an axis of X.
 Parameters
 ----------
 X: ND-Array. Probably should be floats.
 theta (optional): float parameter,used as a multiplier
  prior to exponentiation. Default = 1.0
 axis (optional): axis to compute values along. Default is the
  first non-singleton axis.
 Returns an array the same size as X. The result will sum to 1
 along the specified axis.
 """
 
 # make X at least 2d
 y = np.atleast_2d(X)
 
 # find axis
 if axis is None:
  axis = next(j[0] for j in enumerate(y.shape) if j[1] > 1)
 
 # multiply y against the theta parameter,y = y * float(theta)
 
 # subtract the max for numerical stability
 y = y - np.expand_dims(np.max(y,axis = axis),axis)
 
 # exponentiate y
 y = np.exp(y)
 
 # take the sum along the specified axis
 ax_sum = np.expand_dims(np.sum(y,axis)
 
 # finally: divide elementwise
 p = y / ax_sum
 
 # flatten if X was 1D
 if len(X.shape) == 1: p = p.flatten()
 
 return p
c = np.random.randn(2,3)
print(c)
# 假設第0維是類別，一共有裡兩種類別
cc = softmax(c,axis=0)
# 假設最後一維是類別，一共有3種類別
ccc = softmax(c,axis=-1)
print(cc)
print(ccc)

結果：

c:
[[-1.30022268 0.59127472 1.21384177]
 [ 0.1981082 -0.83686108 -1.54785864]]
cc:
[[0.1826746 0.80661068 0.94057075]
 [0.8173254 0.19338932 0.05942925]]
ccc:
[[0.0500392 0.33172426 0.61823654]
 [0.65371718 0.23222472 0.1140581 ]]

可以看到，對axis=0的軸做softmax時，輸出結果在axis=0軸上和為1(eg: 0.1826746+0.8173254)，同理在axis=1軸上做的話結果的axis=1軸和也為1(eg: 0.0500392+0.33172426+0.61823654)。

這些值是怎麼得到的呢？

以cc為例（沿著axis=0做softmax）：

關於tensorflow softmax函式用法解析

以ccc為例（沿著axis=1做softmax）：

關於tensorflow softmax函式用法解析

知道了計算方法，現在我們再來討論一下這些值的實際意義：

cc[0,0]實際上表示這樣一種概率： P( label = 0 | value = [-1.30022268 0.1981082] = c[*,0] ) = 0.1826746

cc[1,0]實際上表示這樣一種概率： P( label = 1 | value = [-1.30022268 0.1981082] = c[*,0] ) = 0.8173254

ccc[0,0]實際上表示這樣一種概率： P( label = 0 | value = [-1.30022268 0.59127472 1.21384177] = c[0]) = 0.0500392

ccc[0,1]實際上表示這樣一種概率： P( label = 1 | value = [-1.30022268 0.59127472 1.21384177] = c[0]) = 0.33172426

ccc[0,2]實際上表示這樣一種概率： P( label = 2 | value = [-1.30022268 0.59127472 1.21384177] = c[0]) = 0.61823654

將他們擴充套件到更多維的情況：假設c是一個[batch_size,timesteps,categories]的三維tensor

output = tf.nn.softmax(c,axis=-1)

那麼 output[1,2,3] 則表示 P(label =3 | value = c[1,2] )

以上這篇關於tensorflow softmax函式用法解析就是小編分享給大家的全部內容了，希望能給大家一個參考，也希望大家多多支援我們。

關於tensorflow softmax函式用法解析

如下所示： def softmax(logits,axis=None,name=None,dim=None): \"\"\"Computes softmax activations. This function performs the equivalent of

Python lambda表示式filter、map、reduce函式用法解析

前言 lambda是表示式，用於建立匿名函式，可以和filter、map、reduce配合使用。本文環境Python3.7。

JavaScript Array.flat()函式用法解析

在過去的幾年中，已經將許多有用的功能新增到Javascript Array全域性物件中，這些功能為開發人員在編寫可用於陣列的程式碼時提供了多種選擇。這些功能提供了許多優點，其中最值得注意的是，雖然在過去的一段時間裡，

JavaScript立即執行函式用法解析

我們知道，在一般情況下，函式必須先呼叫才能執行，如下所示，我們定義了一個函式，並且呼叫，

Python partial函式原理及用法解析

這篇文章主要介紹了Python partial函式原理及用法解析,文中通過示例程式碼介紹的非常詳細，對大家的學習或者工作具有一定的參考學習價值,需要的朋友可以參考下

python重要函式eval多種用法解析

這篇文章主要介紹了python重要函式eval多種用法解析,文中通過示例程式碼介紹的非常詳細，對大家的學習或者工作具有一定的參考學習價值,需要的朋友可以參考下

JavaScript回撥函式callback用法解析

這篇文章主要介紹了JavaScript回撥函式callback用法解析,文中通過示例程式碼介紹的非常詳細，對大家的學習或者工作具有一定的參考學習價值,需要的朋友可以參考下

Java字串替換函式replace（）用法解析

這篇文章主要介紹了Java字串替換函式replace（）用法解析,文中通過示例程式碼介紹的非常詳細，對大家的學習或者工作具有一定的參考學習價值,需要的朋友可以參考下

關於tf.TFRecordReader()函式的用法解析

讀取tfrecord資料從TFRecords檔案中讀取資料，首先需要用tf.train.string_input_producer生成一個解析佇列。之後呼叫tf.TFRecordReader的tf.parse_single_example解析器。

python統計函式庫scipy.stats的用法解析

背景總結統計工作中幾個常用用法在python統計函式庫scipy.stats的使用範例。正態分佈

php判斷某個方法是否存在函式function_exists (),method_exists()與is_callable()區別與用法解析

本文例項講述了php判斷某個方法是否存在函式function_exists (),method_exists()與is_callable()區別與用法。分享給大家供大家參考，具體如下：

Softmax函式原理及Python實現過程解析

Softmax原理 Softmax函式用於將分類結果歸一化，形成一個概率分佈。作用類似於二分類中的Sigmoid函式。

Python 偏函式用法全方位解析

Python的functools模組中有一種函式叫“偏函式”，自從接觸它以來，發現確實是一個很有用且簡單的函式，相信你看完這篇文章，你也有相見恨晚的感覺。

解析Python 偏函式用法全方位實現

MySql中流程控制函式/統計函式/分組查詢用法解析

路漫漫其修遠兮，吾將上下而求索，又到了週末，我繼續帶各位看官學習回顧Mysql知識。

Python Merge函式原理及用法解析

Merge函式的用法簡單來說Merge函式相當於Excel中的vlookup函式。當我們對2個表進行資料合併的時候需要通過指定兩個表中相同的列作為key，然後通過key匹配到其中要合併在一起的values值。

Python eval函式原理及用法解析

eval函式就是實現list、dict、tuple與str之間的轉化 str函式把list，dict，tuple轉為為字串

tensorflow視覺化——相關函式用法例項

tensorflow版本：1.15.0 tf.summary.scalar，tf.summary.histogram，tf.summary.merge_all，tf.summary.merge，tf.summary.FileWriter，writer.add_summary用法簡單演示！

Python的內建函式，super()超類用法解析

直接用類名呼叫父類方法在使用單繼承的時候沒問題，但是如果使用多繼承，會涉及到查詢順序（MRO）。看到這句話，不太理解,請看這篇文章。

PyTorch中AdaptiveAvgPool函式用法及原理解析

自適應1D池化（AdaptiveAvgPool1d）：對輸入訊號，提供1維的自適應平均池化操作對於任何輸入大小的輸入，可以將輸出尺寸指定為H*W，但是輸入和輸出特徵的數目不會變化。

關於tensorflow softmax函式用法解析

相關推薦