Make money doing the work you believe in
85% cost reduction is believable, but I'd push back a bit on attributing it mainly to the graph engineering vs. Kimi K3's pricing itself — K2 was already priced aggressively below GPT-4 class models, so some of that delta might be model selection, not architecture. Worth isolating the two variables if you haven't: run the same graph setup on a comparably priced model and see how much of the accuracy gain survives.
Aug 4
at
4:15 PM
Relevant people
Log in or sign up
Join the most interesting and insightful discussions.
