Start with disqualifiers: prohibited data use, inability to enforce required approval, unavailable identity controls, no recovery for a consequential action, or an unsupported source-of-truth assumption. An option that fails a nonnegotiable boundary should not recover through a high feature score. Confirm current functionality and commercial availability rather than scoring roadmap intent as delivered capability.
For remaining options, rate the completeness of the end-to-end outcome, source integration, permission granularity, inspectability, exception handling, change management, worker participation, maintenance, and total operating cost. Record confidence and evidence for every rating. A demonstration can show interface behavior, but production reliability, integration depth, and organizational fit require representative testing. Unknown should remain unknown rather than becoming an average score.
Weighting should follow the use case. Recovery and authority may dominate a consequential external action, while simplicity and source fit may matter most for an internal preparation tool. Publish the weights and let responsible stakeholders challenge them before vendor scoring. Otherwise an evaluation team can unconsciously choose the winner by assigning importance after it sees the feature comparison.