1
0 Comments

Switching AI providers cut my per-user costs by 30x

Hey! Quick update on Nokos.

When I first built the product, I used Claude (Anthropic) for all AI features. The problem: I was losing money on every paying user. Plus plan was ¥480/month, but per-user AI costs were ~¥196.

The fix: keep Claude Code for writing the application code (where quality matters most), but switch all production AI features to Gemini Flash (where speed and cost matter more).

Result:
- Chat queries: ¥4.5 → ¥0.07 per call
- Per-user costs dropped 3-7x across all plans
- Break-even went from ~700 users to ~150 users
- Free users can now get real AI features without killing my margins

The lesson: the model that builds your product and the model that runs it don't need to be the same.

Full breakdown:
https://dev.to/tomokiikeda/claude-writes-the-code-gemini-runs-it-how-two-competing-ais-cut-my-saas-costs-by-30x-2hn3

Anyone else running multiple AI providers? How do you decide which model goes where?

posted toAvatar for product Nokos
Nokos