【无标题】

渊博自习室

已于 2023-04-17 21:13:20 修改

阅读量30

点赞数

文章标签： python

于 2023-04-17 21:12:50 首次发布

本文链接：https://blog.csdn.net/m0_61382108/article/details/130209188

版权

找出中文字体个数

图片: Alt

from collections import Counter #导入代码块
import jieba
with open("/Users/mac/Documents/工作/fuel/燃料化验员.txt", "r",encoding="utf-8") as f:
    data = f.read()

    
data = data.translate({ord(c):None for c in list('！| \n/，？℃ , ；、。…《》., ()=;×?-;（）：“”0123456789')})
data = jieba.cut(data)
#for word in data:
    #print(word, '/', end="") #打印出中文分词并使用/分隔

for w, c in Counter(data).most_common():
    print(w, c)