No speed-up in my implementation too

Question

No speed-up in my implementation too

LSC527 opened this issue 2 years ago · comments

I implemented this papaer with torch.autograd.forward_ad. However, fwd gradient showed no speed-up compared to fwd+bwd.

Davide Angioni · Answer 1 · Tue May 24 2022 00:41:01 GMT+0800 (China Standard Time)

It would be interesting for us to see your implementation as well. If you want, you can make a PR to our repo with your code. So we can have multiple implementations available.

Qi Liu · Answer 2 · Thu Jan 18 2024 22:34:28 GMT+0800 (China Standard Time)

I ran the code from the repository, but I couldn't replicate the results mentioned in the paper, especially regarding the CNN. I used the hyperparameter settings specified in the paper.

May I inquire if there are alternative parameter settings available?

Davide Angioni · Answer 3 · Sun Jan 21 2024 00:29:32 GMT+0800 (China Standard Time)

Hi, unfortunately we weren't able to reproduce the same results too.
The hyperparameters we used are the same reported in the paper, but we don't now if alternative hyperparameters settings are available.

We believe that the difference between our implementation and the official one are due to the fact that they did not use functorch