promptdojo_

Mission: baseline model showdown — step 2 of 7

Stage 4 of the mission: you rerun the showdown with a different data seed and the trained rung's accuracy moves from 81% to 76%. What happened?