[学习笔记]人工智能-数据解析和可视化

最新推荐文章于 2022-11-06 17:01:23 发布

法迪

最新推荐文章于 2022-11-06 17:01:23 发布

阅读量714

点赞数

分类专栏： Python基础文章标签： scatter 可视化数据数据解析人工智能 matplotlib.pyplot as plt

本文链接：https://blog.csdn.net/su749520/article/details/79124074

版权

Python基础专栏收录该内容

143 篇文章 3 订阅

订阅专栏

一、投喂学习模型的数据

原始数据

数据可以通过我的github工程下载

https://github.com/sufadi/SimpleNeuronNetworkDemo/blob/master/su/iris.data.csv

1.读取数据

运行示例

# 加载数据原料
import pandas as pd
file = "D:/EclipseProject/PythonStudyBySu/su/iris.data.csv"
# 无文件头
df = pd.read_csv(file, header=None)
# 读取前面 10 行数据
print(df.head(10))

显示结果

     0    1    2    3            4
0  5.1  3.5  1.4  0.2  Iris-setosa
1  4.9  3.0  1.4  0.2  Iris-setosa
2  4.7  3.2  1.3  0.2  Iris-setosa
3  4.6  3.1  1.5  0.2  Iris-setosa
4  5.0  3.6  1.4  0.2  Iris-setosa
5  5.4  3.9  1.7  0.4  Iris-setosa
6  4.6  3.4  1.4  0.3  Iris-setosa
7  5.0  3.4  1.5  0.2  Iris-setosa
8  4.4  2.9  1.4  0.2  Iris-setosa
9  4.9  3.1  1.5  0.1  Iris-setosa

2.数据显示

运行示例

# 数据可视化展示
import matplotlib.pyplot as plt
import numpy as np

y = df.loc[0:100, 4].values
print("显示第四列前100条数据", y)

显示结果

显示第四列前100条数据 ['Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa' 'Iris-setosa'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor' 'Iris-versicolor'
 'Iris-versicolor' 'Iris-versicolor' 'Iris-virginica']

3.对数据进行分类

示例

y = np.where(y == "Iris-setosa", -1, 1)
print("对数据进行分类", y)

运行结果

对数据进行分类 [-1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1
 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1 -1
  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1
  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1  1
  1]

4.抽取出第0和2列的数据

示例

x = df.iloc[0:100, [0, 2]].values
print("抽取出第0和2列的数据", x)

类似一个二维数组

抽取出第0和2列的数据 [[ 5.1  1.4]
 [ 4.9  1.4]
 [ 4.7  1.3]
 [ 4.6  1.5]
 [ 5.   1.4]
 ...........
 [ 6.1  4.6]
 [ 5.8  4. ]
 [ 5.   3.3]
 [ 5.6  4.2]
 [ 5.7  4.2]
 [ 5.7  4.2]
 [ 6.2  4.3]
 [ 5.1  3. ]
 [ 5.7  4.1]]

5 可视化数据

这里的可视化的目的是区分分界线
可视化数据集

# 画出图形
# x 的第一列为x轴，第二列为y轴
plt.scatter(x[:50, 0], x[:50, 1], color='red', marker='o', label='setosa')
plt.scatter(x[50:100, 0], x[50:100, 1], color='blue',
            marker='x', label='versicolor')
plt.xlabel("花瓣长度")
plt.ylabel("花茎长度")
plt.legend(loc='upper left')
plt.show()

6 学习来源

慕课网视频 https://www.imooc.com/video/14379

法迪

关注

0
点赞
踩
1

收藏

觉得还不错? 一键收藏
打赏
0
评论
[学习笔记]人工智能-数据解析和可视化

一、投喂学习模型的数据数据可以通过我的github工程下载https://github.com/sufadi/SimpleNeuronNetworkDemo/blob/master/su/iris.data.csv1.读取数据运行示例# 加载数据原料import pandas as pdfile = "D:/EclipseProject/PythonStudyByS
复制链接

扫一扫