Skip to content
Recommendation freshness

Yesterday's best model can be the wrong choice today.

Model recommendations decay because the model market changes continuously. ModelShortlist is designed to expose current evidence and to tell the host AI when upstream freshness has degraded.

New model launches can change the competitive set immediately.

Price changes can turn an expensive model into the best-value option—or the reverse.

Benchmark and performance evidence can move as independent evaluations change.

Capabilities, context, providers, and ZDR availability can change without the model name changing.

01

Freshness is part of the answer

A model recommendation should not quietly imply “current” when the evidence is stale. ModelShortlist tracks upstream source freshness independently so the host AI can distinguish fresh, stale, and unavailable evidence.

02

Different upstream failures mean different things

Artificial Analysis, the OpenRouter model catalog, and OpenRouter ZDR endpoint evidence are treated as separate sources rather than one all-or-nothing dependency.

If Artificial Analysis is temporarily unavailable, OpenRouter models can remain eligible with missing benchmark evidence.

If ZDR endpoint evidence is unavailable and ZDR was not requested, the general model catalog can still be used.

If ZDR is required but current or cached ZDR evidence is unavailable, ModelShortlist fails closed rather than guessing.

If the required OpenRouter model catalog is unavailable with no usable cache, ModelShortlist does not fabricate a recommendation.

03

Stale evidence is labeled, not disguised

A last-known-good source can be useful during a temporary outage, but only if its age and degraded state are explicit. ModelShortlist surfaces that state in tool output so the host AI can qualify the recommendation appropriately.

Try it in your assistant

Turn the concept into a current decision.

These prompts ask for workload-specific reasoning rather than a permanent ranking. The answer can change as benchmark and operational evidence changes.

Recommend models for this workload, but tell me whether each upstream evidence source is fresh, stale, or unavailable.
If Artificial Analysis is temporarily unavailable, give me the best operational shortlist you can without pretending the benchmark evidence is current.
I require ZDR. If current ZDR endpoint evidence is unavailable, do not infer or guess eligibility.

Ask the current model market, not yesterday's blog post.

ModelShortlist runs locally, uses your Artificial Analysis and OpenRouter keys, and gives your host AI current evidence for the workload you actually care about.

Open the install configurator