No bio yet.
OneTriangle optimizes multi-model inference by transferring KV cache between models, cutting redundant prefill cost.