The final shortlist should not be formula-only.
A participant can have a good score and still be the wrong fit because the behavior is obviously unscalable, excessively noisy, too concentrated, dependent on simulator artifacts, or simply inconsistent with the kind of PM Finam would want to back.
That discretion should be explicit from the beginning.