I use PI WEB (<a href="https://pi-web.dev" rel="nofollow">https://pi-web.dev, there's more than one with that name) to orchestrate remotely. Pi lets me run local models (currently Qwen 3.8 Flash Next on Strix Halo 128GB) and flagship models (with subscription auth, not API pricing) side by side. I typically ask GPT 6.1 Sol to review requirements and then spawn a subsession with Qwen, review Qwen's work, ask Qwen to fix. About 95% success with a single pass like that, achieving flagship quality without the price tag (albeit slower).
I started using Pi because it has a small system prompt and local models were too slow to start. Then I started adding custom skills and extensions when I hit little corner cases. It's been so easy to bend into what I need.
wasting_time · · focus · HN ↗
pettijohn · · focus · HN ↗
I started using Pi because it has a small system prompt and local models were too slow to start. Then I started adding custom skills and extensions when I hit little corner cases. It's been so easy to bend into what I need.