Machine Learning A-Z学习笔记6

最新推荐文章于 2024-07-23 20:25:48 发布

JiduoW

最新推荐文章于 2024-07-23 20:25:48 发布

阅读量111

点赞数

分类专栏：机器学习笔记文章标签：机器学习学习 python

本文链接：https://blog.csdn.net/JiduoW/article/details/123262350

版权

机器学习笔记专栏收录该内容

17 篇文章 18 订阅

订阅专栏

Machine Learning A-Z学习笔记6

第六章逻辑回归

1.简单原理

如以下例子，运用线性回归不能很好的拟合模型

在这里插入图片描述

此时可以采用逻辑预测购买者的购买概率

在这里插入图片描述

通过Sigmoid函数将元贝的回归线变成范围在0~1之间的曲线，代表0%购买率和100%购买率

在这里插入图片描述

假设年龄为是20、30、40、50，则相应的购买概率为0.7%、23%、85%、99.4%。

在这里插入图片描述

预测购买率大于50%的则认为会进行购买

在这里插入图片描述

2.相关代码


# Logistic Regression
"""
逻辑回归:年紀+收入vs购买车辆
"""

# Importing the libraries

import numpy as np
import matplotlib.pyplot as plt
import pandas as pd

# Importing the dataset

dataset = pd.read_csv('Social_Network_Ads.csv')
X = dataset.iloc[:, [2,3]].values
y = dataset.iloc[:, 4].values

# 数据分割
from sklearn.model_selection import train_test_split
X_train, X_test, y_train, y_test = train_test_split(X, y, test_size = 0.25, random_state = 0)

# 特征缩放
from sklearn.preprocessing import StandardScaler
sc_X = StandardScaler()
X_train = sc_X.fit_transform(X_train)
X_test = sc_X.transform(X_test)

# 模型训练-random_state:打乱数据的顺序

from sklearn.linear_model import LogisticRegression
classifier = LogisticRegression(random_state = 0)
classifier.fit(X_train, y_train)

# Predicting the Test set results

y_pred = classifier.predict(X_test)

# 混淆矩阵
from sklearn.metrics import confusion_matrix
cm = confusion_matrix(y_test, y_pred)

# 训练结果可视化

from matplotlib.colors import ListedColormap

X_set, y_set = X_train, y_train
X1, X2 = np.meshgrid(np.arange(start = X_set[:, 0].min() - 1, stop = X_set[:, 0].max() + 1, step = 0.01),
                     np.arange(start = X_set[:, 1].min() - 1, stop = X_set[:, 1].max() + 1, step = 0.01))
plt.contourf(X1, X2, classifier.predict(np.array([X1.ravel(), X2.ravel()]).T).reshape(X1.shape),
             alpha = 0.75, cmap = ListedColormap(('red', 'green')))
plt.plot(X2)
plt.xlim(X1.min(), X1.max())
plt.ylim(X2.min(), X2.max())
for i, j in enumerate(np.unique(y_set)):
    plt.scatter(X_set[y_set == j, 0], X_set[y_set == j, 1],
                c = ListedColormap(('orange', 'blue'))(i), label = j)
plt.title('Logistic Regression (Training set)')
plt.xlabel('Age')
plt.ylabel('Estimated Salary')
plt.legend()
plt.show()

# Visualising the Test set results
from matplotlib.colors import ListedColormap
X_set, y_set = X_test, y_test
X1, X2 = np.meshgrid(np.arange(start = X_set[:, 0].min() - 1, stop = X_set[:, 0].max() + 1, step = 0.01),
                      np.arange(start = X_set[:, 1].min() - 1, stop = X_set[:, 1].max() + 1, step = 0.01))
plt.contourf(X1, X2, classifier.predict(np.array([X1.ravel(), X2.ravel()]).T).reshape(X1.shape),
              alpha = 0.75, cmap = ListedColormap(('red', 'green')))
plt.xlim(X1.min(), X1.max())
plt.ylim(X2.min(), X2.max())
for i, j in enumerate(np.unique(y_set)):
    plt.scatter(X_set[y_set == j, 0], X_set[y_set == j, 1],
                c = ListedColormap(('orange', 'blue'))(i), label = j)
plt.title('Logistic Regression (Test set)')
plt.xlabel('Age')
plt.ylabel('Estimated Salary')
plt.legend()
plt.show()

JiduoW

关注

0
点赞
踩
0

收藏

觉得还不错? 一键收藏
0
评论
Machine Learning A-Z学习笔记6

Machine Learning A-Z学习笔记6第六章逻辑回归1.简单原理如以下例子，运用线性回归不能很好的拟合模型此时可以采用逻辑预测购买者的购买概率通过Sigmoid函数将元贝的回归线变成范围在0~1之间的曲线，代表0%购买率和100%购买率假设年龄为是20、30、40、50，则相应的购买概率为0.7%、23%、85%、99.4%。预测购买率大于50%的则认为会进行购买2.相关代码# Logistic Regression"""逻辑回归:年紀+收入vs购买车辆"""
复制链接

扫一扫