Python数据建模--分类

最新推荐文章于 2024-02-29 16:58:25 发布

小白-小天

最新推荐文章于 2024-02-29 16:58:25 发布

阅读量1.9k

点赞数 3

分类专栏： Python 数据分析数据建模

本文链接：https://blog.csdn.net/qq_42169061/article/details/106135795

版权

Python 同时被 3 个专栏收录

53 篇文章 12 订阅

订阅专栏

数据分析

44 篇文章 4 订阅

订阅专栏

数据建模

8 篇文章 3 订阅

订阅专栏

电影分类

导入库

import numpy as np
import pandas as pd
import matplotlib.pyplot as plt
%matplotlib inline

数据创建

from sklearn import neighbors  # 导入KNN分类模块
import warnings
warnings.filterwarnings('ignore') 
# 不发出警告

data = pd.DataFrame({'name':['北京遇上西雅图','喜欢你','疯狂动物城','战狼2','力王','敢死队'],
                  'fight':[3,2,1,101,99,98],
                  'kiss':[104,100,81,10,5,2],
                  'type':['Romance','Romance','Romance','Action','Action','Action']})
print(data)
print('-------')

创建knn模型，并预测【18,90】

knn = neighbors.KNeighborsClassifier()   # 取得knn分类器
knn.fit(data[['fight','kiss']], data['type'])
print('预测电影类型为:', knn.predict([[18, 90]]))
# 加载数据，构建KNN分类模型
# 预测未知数据

在图中展示各电影位置

plt.scatter(data[data['type'] == 'Romance']['fight'],data[data['type'] == 'Romance']['kiss'],color = 'r',marker = 'o',label = 'Romance')
plt.scatter(data[data['type'] == 'Action']['fight'],data[data['type'] == 'Action']['kiss'],color = 'g',marker = 'o',label = 'Action')
plt.grid()
plt.legend()
plt.scatter(18,90,color = 'r',marker = 'x',label = 'Romance')
plt.ylabel('kiss')
plt.xlabel('fight')
plt.text(18,90,'《你的名字》',color = 'r')

增加数据量进行模型训练

data2 = pd.DataFrame(np.random.randn(100,2)*50,columns = ['fight','kiss'])
data2['type'] = knn.predict(data2)
print(data2.head())

图中展示


plt.scatter(data[data['type'] == 'Romance']['fight'],data[data['type'] == 'Romance']['kiss'],color = 'r',marker = 'o',label = 'Romance')
plt.scatter(data[data['type'] == 'Action']['fight'],data[data['type'] == 'Action']['kiss'],color = 'g',marker = 'o',label = 'Action')
plt.grid()
plt.scatter(data2[data2['type'] == 'Romance']['fight'],data2[data2['type'] == 'Romance']['kiss'],color = 'r',marker = 'x',label = 'Romance')
plt.scatter(data2[data2['type'] == 'Action']['fight'],data2[data2['type'] == 'Action']['kiss'],color = 'g',marker = 'x',label = 'Action')
plt.legend()
plt.ylabel('kiss')
plt.xlabel('fight')

植物分类

数据导入并输出数据特征

from sklearn import datasets
iris = datasets.load_iris()
print(iris.keys())
print('数据长度为:%i条' % len(iris['data']))

print(iris.feature_names)
print(iris.target_names)
#print(iris.target)
print(iris.data[:5])
# 150个实例数据
# feature_names - 特征分类：萼片长度，萼片宽度，花瓣长度，花瓣宽度  → sepal length, sepal width, petal length, petal width
# 目标类别：Iris setosa, Iris versicolor, Iris virginica.

把数字转换为标记名字

df = pd.DataFrame(iris.data, columns = iris.feature_names)  # 将特征值转为Dataframe
df['target'] = iris.target 
ty = pd.DataFrame({'target':[0,1,2],
                  'target_names':iris.target_names})
df = pd.merge(df, ty, on = 'target')
# 数据转换

训练模型并预测

knn = neighbors.KNeighborsClassifier()   # 取得knn分类器
knn.fit(iris.data, df['target_names'])
# 建立分类模型

pre_data = [[0.1, 0.2, 0.3, 0.4]]
print('预测结果为:', knn.predict(pre_data))
# 预测结果

df.head()
# 显示数据

Python 数据建模：

- Python数据建模–回归
 - Python数据建模–分类
 - Python数据建模–主成分分析
 - Python数据建模–K-means聚类
 - Python数据建模–蒙特卡洛模拟

小白-小天

关注

3
点赞
踩
10

收藏

觉得还不错? 一键收藏
2
评论
Python数据建模--分类

分类电影分类导入库数据创建创建knn模型，并预测【18,90】在图中展示各电影位置增加数据量进行模型训练图中展示植物分类数据导入并输出数据特征把数字转换为标记名字训练模型并预测最邻近分类的python实现方法介绍：在距离空间里，如果一个样本的最接近的k个邻居里，绝大多数属于某个类别，则该样本也属于这个类别实例：电影分类 / 植物分类电影分类导入库import numpy as npimport pandas as pdimport matplotlib.pyplot as plt%matp
复制链接

扫一扫