爬虫基础（笔记+一点点代码）

最新推荐文章于 2023-06-20 17:39:18 发布

qq_45776928

最新推荐文章于 2023-06-20 17:39:18 发布

阅读量96

点赞数

分类专栏：笔记

本文链接：https://blog.csdn.net/qq_45776928/article/details/104985996

版权

笔记专栏收录该内容

1 篇文章 0 订阅

订阅专栏

标题：爬虫基础（笔记+一点点代码）

在这里插入图片描述

import  urllib.request

def load_data():
    url = "http://www.baidu.com/"
    #get的请求
    #http请求
    #response：http相应的对象
    response = urllib.request.urlopen(url)
    print(response)
    #读取内容  bytes类型
    data = response.read()
    print(data)
    #将文件获取的内容转换成字符串
    str_data = data.decode("utf-8")
    print(str_data)
    #将数据写入文件
    with open("baidu.html","w",encoding="utf-8")as f:
        f.write(str_data)
    #将字符串类型转化成bytes
    str_name = "baidu"
    bytes_name = str_name.encode("utf-8")
    print(bytes_name)

    #python派出户的类型：str bytes
    #如果爬取回来的是bytes类型：但是你写的时候需要字符串 decode("utf-8")
    #如果爬取过来的是str类型：但是你要写入的是bytes类型，encode("utf-8")
load_data()

qq_45776928

关注

0
点赞
踩
0

收藏

觉得还不错? 一键收藏
0
评论
爬虫基础（笔记+一点点代码）

标题：爬虫基础（笔记+一点点代码）import urllib.requestdef load_data(): url = "http://www.baidu.com/" #get的请求 #http请求 #response：http相应的对象 response = urllib.request.urlopen(url) print(...
复制链接

扫一扫