AI Model Shortlist Wizard
Create a starting shortlist from the verified catalog. This is a transparent rule-based aid, not a sponsored ranking or a prediction of your monthly bill.
Data snapshot: . Final selection should include your own quality, latency, privacy, and reliability tests.
Your shortlist
Rates are USD per one million input/output tokens. Temporary pricing and cache assumptions are shown where applicable. Use the cost calculator after estimating your actual workload.
What this wizard does
The wizard applies visible rules to a deliberately small reviewed catalog. Priority selects one of three maintained candidate groups; workload can add models associated with complex, multimodal, or high-volume use; and the context choice removes models below the requested window. The result is a shortlist, not a winner.
Cost priority
Favors lower published input/output rates. It does not know your prompt length, cache-hit rate, retries, or quality threshold.
Balanced priority
Includes models positioned between unit cost and broad capability. “Balanced” is an editorial rule, not a scientific score.
Capability priority
Starts with models intended for more demanding work. You still need task-specific evidence before deployment.
How to validate the shortlist
- Write an acceptance test for the actual task.
- Run the same representative examples through every candidate.
- Record task success, critical failures, latency, token usage, retries, and tool calls.
- Reject candidates that fail privacy or safety gates, regardless of their weighted score.
- Choose by cost per accepted outcome and keep a rollback option.
Important limitations
- The catalog is selective and does not represent every provider or deployment option.
- Provider features, regions, rate limits, and prices can change after the displayed verification date.
- Published context size does not guarantee reliable performance throughout the full window.
- No vendor pays for placement or receives a hidden scoring advantage.