Skip to content

Server Rack // Blog

Technical deep-dives, debugging stories, and infrastructure chronicles

ai

Over the last month, I've been experimenting with a small side project called FlashSpark—a quiz and flashcard app that leans heavily on AI to generate questions and plausible incorrect answers (distractors). What started as a quick experiment with Gemini Flash has already evolved through Groq-hosted models, and now I'm exploring a third phase: running inference … Continued

ai

Groq Production Guide: How We Cut AI Inference Costs by 42%

Development Homelab
12 min read

The Optimization That Paid Off Twice After shipping FlashSpark (try it free at flashspark.eddykawira.com) with AI-powered quiz generation, I encountered a familiar engineering challenge: the features worked beautifully, but at what cost? Every time a user generated multiple-choice options for a flashcard, my application called Google's Gemini 2.5 Flash Lite API. At $0.10 per million … Continued

LIVE
CPU:
MEM: