Refacto Agents

Industry story

AWS P6-B300 GPU instances expand to Seoul Region

cloud-costs gpu-supply inference

Seoul getting AWS P6-B300 instances matters more than a routine regional rollout. These are the serious boxes: 8 NVIDIA Blackwell Ultra GPUs per instance, 2.1 TB of GPU memory, and 6.4 Tbps of interconnect, which is 2x the bandwidth of the previous generation. For teams training or serving trillion-parameter models in Asia Pacific, that memory ceiling and networking headroom are the constraints that actually bind. Watch whether Korean and Japanese foundation model builders start consolidating workloads on AWS rather than building out their own clusters.

Full analysis

Amazon Web Services has made its EC2 P6-B300 instances available in the Asia Pacific (Seoul) Region, adding to existing deployments in Oregon, GovCloud US-East, and N. Virginia. Each instance packs 8 NVIDIA Blackwell Ultra GPUs with 2.1 TB of high-bandwidth GPU memory, 6.4 Tbps EFA (Elastic Fabric Adapter, a high-speed interconnect for GPU clusters) networking, and 4 TB of system memory. Compared to the previous P6-B200 generation, the B300 instances deliver 2× networking bandwidth, 1.5× GPU memory, and 1.5× compute throughput (measured in TFLOPS at FP4 precision), making them better suited for training and serving large trillion-parameter foundation models. Higher memory capacity and faster inter-node networking directly benefit agent deployments that rely on very large language models, reducing training times and increasing token throughput.

Comments