The Math That Makes It Worth It
ChatGPT Plus costs $20/month. Over a year: $240. Over two years: $480.
A used mini PC that runs local AI models costs about $351. It pays for itself in 18 months vs ChatGPT Plus alone. If you also replace Midjourney ($10/mo) and a TTS subscription ($5/mo), the payoff drops to under a year.
And after that, it’s free. Forever. No subscription. No price hike. No account suspension.
The Build
What You Need
| Component | What | Price |
|---|---|---|
| Mini PC | Used Lenovo ThinkCentre Tiny (M720q or M920q) | $120-180 |
| RAM | 32GB DDR4 SODIMM (upgrade from stock 8GB) | $40-60 |
| Storage | 1TB NVMe SSD | $50-70 |
| Total | $210-310 |
Look for deals on eBay, Facebook Marketplace, or Amazon Renewed. The Lenovo Tiny series is ideal because it’s tiny (1L chassis), silent, and supports up to 32GB DDR4.
Alternatives
| Mini PC | Price | RAM | Notes |
|---|---|---|---|
| Beelink S12 Pro | $200 | 16GB | N100 CPU, budget pick |
| Lenovo M720q | $150 | 32GB | Best value, upgradable |
| Dell Optiplex Micro | $140 | 16GB | Widely available used |
| Minisforum UN100L | $250 | 16GB | New, low power |
Software Setup (15 Minutes)
Step 1: Install Linux (or stay on Windows)
Ubuntu 24.04 LTS is recommended — better GPU support, Docker is smoother, and most AI tools assume Linux.
But Windows works too. Ollama runs natively on Windows now.
Step 2: Install Ollama
| |
Step 3: Pull Your First Model
| |
Step 4: Add a Web UI
| |
Open http://localhost:3000. You now have a ChatGPT replacement running on a $351 box in your house.
What You Get For $351
- ChatGPT replacement — unlimited messages, no rate limits
- Privacy — nothing leaves your network
- Offline capability — works without internet
- No account — no email, no phone number, no tracking
- Customizable — switch models anytime, use uncensored models
- API access — OpenAI-compatible API on localhost:11434 for automation
What You Don’t Get
- GPT-4 level reasoning — local 7B-8B models are good but not at that level. The gap is closing fast.
- Image generation — you need a GPU for that. This build is CPU-only.
- Cloud sync — your chats live on this box. Back them up.
Adding a GPU Later (Optional)
Want image generation and faster inference? Add a used GPU:
| GPU | Price | VRAM | What It Unlocks |
|---|---|---|---|
| RTX 3060 Ti | $150-200 | 8GB | 7B models at speed, image generation |
| RTX 3080 10GB | $250-350 | 10GB | 9B-13B models, faster everything |
| RTX 3090 | $700-900 | 24GB | 30B models, serious AI workstation |
You don’t need to buy these new. eBay and r/hardwareswap have used cards daily.
Real-World Performance
On a Lenovo M720q with 32GB RAM, no GPU:
| Model | Size | Load Time | Tokens/sec |
|---|---|---|---|
| llama3.2:3b | 3B | 3s | 25-40 t/s |
| llama3.1:8b | 8B | 5s | 8-15 t/s |
| qwen2.5:7b | 7B | 4s | 10-18 t/s |
That’s fast enough for real conversation. 10 tokens/second reads faster than you can type.
The Real Cost Comparison
| Setup | Year 1 | Year 2 | Year 3 |
|---|---|---|---|
| ChatGPT Plus only | $240 | $480 | $720 |
| ChatGPT + Midjourney + TTS | $420 | $840 | $1,260 |
| Mini PC self-hosted | $351 | $351 | $351 |
| Savings (vs full stack) | $69 | $489 | $909 |
After year 1, it’s all savings. The mini PC keeps running. The subscriptions keep charging.
Ready to build? Check the VRAM Calculator to see what models your hardware can run.