DeepSeek V4-Pro served on W&B Inference (CoreWeave). Stronger reasoning and coding than V4-Flash at $1.74/$3.46 per 1M tokens — a solid step up for tasks that need more than the cheap tier but don't require a frontier proprietary model.
LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.