Topic · unordered discovery
Benchmarking
Latency, throughput, resource use, and fair performance measurement.
Engineering
Writing
模型換了執行方式,DQN 還會做出相同選擇嗎?
把訓練好的打磚塊模型交給另一種執行環境前,先確認它面對相同畫面時還會做出相同選擇,再比較速度與代價。
Reinforcement LearningONNXModel Inference
訓練更快,DQN 就更好嗎?
同時跑更多場 Breakout 確實能加快經驗收集,但速度不等於學得更好。從模型比較與新局面測試,看一次高分能說明多少。
Reinforcement LearningDQNEvaluation