provider::kimi::abort
Cancel the in-flight upstream stream for a request_id (router::abort fan-out), stopping billed generation immediately.
- request_idstringrequired
- abortedbooleanrequired
Moonshot (Kimi) Chat Completions provider worker; implements provider::kimi::stream and provider::kimi::refresh_models behind llm-router.
exact versions are immutable; binary and bundle artifacts are digest-pinned.
Cancel the in-flight upstream stream for a request_id (router::abort fan-out), stopping billed generation immediately.
Count prompt tokens for {model, system_prompt?, tools?, messages} through Moonshot's own estimator endpoint; never runs the model and costs nothing.
Wire agent messages, the same shape `provider::kimi::stream` accepts. Must be non-empty.
Single-variant role tags: exact-match on deserialize, correct wire string on serialize, and they let `AgentMessage` be an untagged union.
Model id the prompt targets (required by the upstream estimator).
System prompt counted as the leading wire message when present.
Function invocation schemas; mapped to the wire `tools` array.
Always `metered`: the count came from the upstream itself.
Prompt tokens the upstream estimator counted for the assembled request.
Internal: router::ready subscriber that re-declares this provider and refreshes its catalog.
Refresh the Kimi catalog slice from GET /v1/models and reconcile it through the router; returns the model count written.
Stream a Kimi (Moonshot) chat completion: resolve credentials, call the upstream Chat Completions API, and relay AssistantMessageEvent frames to writer_ref.
Single-variant role tags: exact-match on deserialize, correct wire string on serialize, and they let `AgentMessage` be an untagged union.