Hey! Quick update on Nokos.
When I first built the product, I used Claude (Anthropic) for all AI features. The problem: I was losing money on every paying user. Plus plan was ¥480/month, but per-user AI costs were ~¥196.
The fix: keep Claude Code for writing the application code (where quality matters most), but switch all production AI features to Gemini Flash (where speed and cost matter more).
Result:
- Chat queries: ¥4.5 → ¥0.07 per call
- Per-user costs dropped 3-7x across all plans
- Break-even went from ~700 users to ~150 users
- Free users can now get real AI features without killing my margins
The lesson: the model that builds your product and the model that runs it don't need to be the same.
Full breakdown:
https://dev.to/tomokiikeda/claude-writes-the-code-gemini-runs-it-how-two-competing-ais-cut-my-saas-costs-by-30x-2hn3
Anyone else running multiple AI providers? How do you decide which model goes where?