Kimi K3 via Baseten
consult_kimi3_baseten is a distinct, explicitly approved, metered second-opinion route through PorkiCoder's authenticated proxy.
PorkiCoder names every Baseten-backed path so users know where paid work runs. Use the distinct Kimi K3 consult for a Baseten-hosted second opinion, or use the existing managed fan-out and HogCode routes for their supported Kimi and GLM workflows.
consult_kimi3_baseten is a distinct, explicitly approved, metered second-opinion route through PorkiCoder's authenticated proxy.
Kimi K2.7 Code and GLM-5.2 can run as isolated fan-out candidates. GLM-5.2 can judge results, while HogCode uses its managed Kimi K2.7 Code route.
consult_kimi, consult_kimi27, consult_kimi27_fast, and consult_glm52 remain retired. They are not aliases for the new Kimi K3 route.
The Baseten-backed paths are wired into PorkiCoder's managed workflows. Choose the named route you intend to authorize; PorkiCoder never substitutes one provider for another.
consult_kimi3_baseten, run Kimi K2.7 Code or GLM-5.2 in isolated worktrees, or launch hogcode.Each provider path is named and scoped so users can tell where a paid request runs.
| Route | Transport | Purpose |
|---|---|---|
| Kimi K2.7 Code | PorkiCoder-managed Baseten routing for fan-out and HogCode. | Isolated implementation or review candidates; read-only project assistance in HogCode. |
| GLM-5.2 | PorkiCoder-managed Baseten routing for fan-out workers and judges. | Large-context candidate reviews, architecture tradeoffs, and fan-out judging. |
Kimi K3 via Basetenconsult_kimi3_baseten |
Authenticated PorkiCoder proxy; the production Baseten credential remains server-side. | An explicitly approved, metered Kimi K3 second opinion with clear Baseten attribution. |
consult_kimi3 uses Kimi Platform; consult_kimi3_baseten uses Baseten. Neither silently fails over to the other.As checked on August 10, 2026, Baseten publishes default Model API limits of 120 requests per minute and 1,000,000 tokens per minute for Pro. Verified Basic and Pro publish the same default 120 RPM, so dependable scale still requires workload-appropriate sustained and burst capacity. PorkiCoder queues shared work and handles explicit rate-limit failures, but does not multiply quotas with extra keys.
Baseten's published Kimi K3 Model API rates are $3.00 per million uncached input tokens, $0.30 per million cached input tokens, and $15.00 per million output tokens. Provider pricing and limits can change; the official pages remain the source of truth.
Name the route and authorization in plain English:
consult_kimi3_baseten to review this architecture.