v0.13.3 — merge ollama pullers into single service

- One ollama-pull service with PULL_CPU/PULL_CUDA flags replaces
  separate ollama-pull + ollama-cuda-pull
- CUDA puller now pulls all models (CPU + CUDA)
- Removed stale ollama-cuda-pull refs from recommend-limits.sh
- Service count 28 → 27