The model is only part of the story.
A model’s origin and the location of the service running it are different questions. An open-weight model can be hosted by another company; a seemingly local service can also forward requests to an upstream API.
I wanted to evaluate the actual inference path: the provider, its infrastructure, its contractual commitments, and where a prompt might end up.