python手写kmeans算法

最新推荐文章于 2022-04-29 15:30:09 发布

置顶

菜鸟懿

最新推荐文章于 2022-04-29 15:30:09 发布

阅读量668

点赞数

CC 4.0 BY-SA版权

分类专栏：机器学习文章标签：聚类算法 python

本文链接：https://blog.youkuaiyun.com/coder_yi_liu/article/details/112535466

本文详细介绍了如何从零开始用Python实现KMeans聚类算法，内容包括算法的基本原理及步骤，通过实例展示了如何在没有依赖库的情况下完成聚类过程。

摘要生成于 C知道，由 DeepSeek-R1 满血版支持，前往体验 >

kmean聚类是最基础和常见的算法，工程上使用比较常见，spark, sklearn都有实现，本文手写实现kmeans

#!/usr/bin/python
import sys
import random
import math

def create_rand_points(max_x, max_y, count):
    """Create count points (0-x), (0-y).

    """
    points = []
    for i in range(0, count):
        x = random.randint(0, max_x)
        y = random.randint(0, max_y)
        points.append([x,y])
    return points

def get_start_k_points(points, k):
    """Get k start points.

    """
    if k > len(points):
        return None
    random.shuffle(points)
    return points[0:k]

def get_nearest_point_index(point, central_points):
    """
    """
    min_dis = 2000000000
    index = 0
    for i in range(0, len(central_points)):
        dis = 0.0
        for j in range(0, len(point)):
            dis += (point[j]-central_points[i][j]) * (point[j]-central_points[i][j])
        if mat

最低0.47元/天解锁文章