Beyond the platform, the models are callable directly over the API — the way you would call any other model provider. Speech-to-text arrives first; text-to-speech follows as coming soon.
One runtime underneath all of it. The engine is what the platform runs on; the models are the layers it owns, arriving on their own as they are ready.
Model pricing is published once go-to-market pricing is established. It is listed in the navigation and marked coming soon until then.