About calculation of loss function

Hi, I have some questions about this loss function because there are some differences between the paper and the code.
First of all, the paper:
<img width="1267" alt="Screen Shot 2021-11-27 at 19 26 52" src="https://user-images.githubusercontent.com/51821219/143679239-8ec54dbb-8938-4229-abc5-783069cf9b46.png">
The loss is the sum of the average logsoftmax between the query point and other prototypes and the average distance between the query point and the corresponding prototypes.

But in the code:
``` python
       log_p_y = F.log_softmax(-dists, dim=1).view(n_class, n_query, -1)
       loss_val = -log_p_y.gather(2, target_inds).squeeze().view(-1).mean()
```
The loss in code is **only** the sum of LogSoftmax  between the query point and the corresponding prototypes.

I'm confused. Is it my understanding of the code or my understanding of the paper?

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

About calculation of loss function #28

Metadata

Assignees

Labels

Projects

Milestone

Relationships

Development

About calculation of loss function #28

Description

Metadata

Metadata

Assignees

Labels

Projects

Milestone

Relationships

Development

Issue actions