he spent $1,999 once. saved $21,000 in a year. NVIDIA DGX Spark on his desk
was paying $180/month for cloud GPUs. every experiment had a price tag. every overnight run cost money. then he bought the machine
128GB unified memory. 1 petaflop at FP4. runs Qwen, Llama, Mistral locally. nothing leaves the machine. ever
stack: Ollama. PyTorch. HuggingFace. vLLM. all compatible out of the box
cloud cost per year: $2,160+. DGX Spark: $1,999 once. electricity after that: basically zero
Jensen Huang hand-delivered the first units to Elon Musk at Starbase. fits next to a monitor. no rack. no server room. no cooling setup
the math only works once. after that every inference run is free
显示更多