Draw by hand
n=5 completions are sampled for one task. Toggle each sample between pass and fail, drag the budget k, and watch the plug-in, the unbiased estimator, and a full brute-force enumeration update together.
All C(5,2) = 10 size-2 subsets
Why plug-in lies
Fix n=5, k=2, and a true per-sample pass rate p. Slide p and compare the exact truth, 1-(1-p)^k, against the exact expectation of each estimator over c ~ Binomial(5, p), no sampling involved.