A convolutional backbone follows the usual convention: every time pooling halves the grid, the filter count doubles. Two of its convolutional layers, both with "same" padding, sit in consecutive stages:
| layer | input block | filters | output block |
|---|---|---|---|
| P | |||
| Q |
For each layer, count two different things for a single image:
Which statement about compared with is correct?
Select all that apply.