3 pointsby stared8 hours ago1 comment
  • mrrrcs6 hours ago
    One run per cell is more noise than the gap you are measuring
    • stared3 minutes ago
      These was some code budget there.

      But as from seeing various runs, errors bars are gross overestimation (as not "the same test", but "if we have different tasks from the same sample").