Sizing bands
| Agent pattern | RAM starting point | Notes |
|---|---|---|
| API-only worker (OpenAI/Anthropic tools) | 2–4GB | Node/Python + queue; watch browser tools |
| RAG + embeddings DB on box | 8–16GB | Vector DB + app memory dominate |
| Headless browser / computer-use agent | 8GB+ | Chromium tabs spike RSS hard |
| Local 7B quantized model | 16GB+ | More if concurrent agents share weights |
| Larger local models / GPU path | GPU host | RAM still needed for KV + app; VPS GPU SKUs vary |
Where to buy the RAM (scored providers)
Value and performance scores help when you intentionally oversize memory.
Ops tips that save RAM
Related: Do you need root for AI agents? and Best VPS for AI agents.
Frequently Asked Questions
See Also
Best VPS for AI agents · Root access for AI agents · Docker & homelab · UpCloud vs Hetzner