The pricing alone! RTX 6000 Pro at $0.66/hr when runPod was quoting me $2.09 for the same GPU. Spun up in under 5 mins with no weird credits nonsense, just pay for what you use. Been running inference workloads on it consistently, performance has been solid throughout
Hey Product Hunt! 👋
We built packet.ai because we kept seeing the same problem — developers and ML teams overpaying for GPU compute, dealing with spot instance interruptions killing training runs, and being forced into 8x GPU configs on AWS when they only needed one.
hosted.ai has been running infrastructure since 1996. We know what reliable compute feels like. packet.ai is our answer to a market that's been overcharging and underdelivering for too long.
Here's what we launched with:
🔹 RTX Pro 6000 Blackwell (96GB) - from $0.66/hr
🔹 L40S (48GB) - from $0.60/hr
🔹 A100 PCIe - $1.43/hr (single GPU - no forced 8x)
🔹 B200 (180GB HBM3e) - $3.75/hr Dynamic / $5.9/hr Dedicated (Hourly & monthly available)
All on-demand. No spot. No contracts. No hidden fees. Deploy in under 5 minutes.
One-click deploys for vLLM, Jupyter, ComfyUI, HuggingFace TGI, VS Code, Langflow and more — so you're not spending half a day setting up your environment.
Would love your feedback — especially curious what GPU workloads you're running and what's been frustrating about your current setup. Happy to answer anything!
Spun up a B200 in about four minutes and was SSHed in before my coffee got cold, pricing really is noticeably better than AWS for the long runs.
Being able to save a custom config of installed drivers and CUDA versions would save a ton of time on repeat deploys, especially when spinning up multiple identical clusters for experiments.
Have been waiting for a solid European GPU option and the pricing looks legit. One thing that would push me over the edge is a built-in benchmark suite so I can see real-world throughput per dollar before committing to a reservation, instead of guessing based on the GPU model alone.
the pricing looks solid for the hardware you get. one thing that would actually help me decide faster is a live availability calendar per gpu type so i can see when slots open up without having to ping support. right now it's kind of a guessing game on whether i'll get a b200 today or next week
the 5-minute SSH deploy promise is genuinely the kind of detail that shows the team actually uses their own product instead of just marketing it