Please kindly disclose all the hyper-parameters you tuned, and how much you tuned for each of them. Example form: "learning rate: [1e-5*, 1e-4, 1e-3], num_layers: [1, 2, 3*, 4, 5], ...", where the asterisks denote the hyper-parameters you eventually selected to report the test performance. This information will not appear in the leaderboard for the time being, but it is important for us to keep the record and encourage the fair model comparison.