python实现统计文本中单词出现的频率详解

时间：2021-06-18 08:45:07|栏目：Python代码|点击：次

本文实例为大家分享了python统计文本中单词出现频率的具体代码，供大家参考，具体内容如下

#coding=utf-8
import os
from collections import Counter
sumsdata=[]
for fname in os.listdir(os.getcwd()):
  if os.path.isfile(fname) and fname.endswith('.txt'):
    with open(fname,'r') as fp:
      data=fp.readlines()
    sumsdata+=[line.strip().lower() for line in data]
cnt=Counter()
for word in sumsdata:
  cnt[word]+=1
cnt=dict(cnt)
for key,value in cnt.items():
  print(key+":"+str(value))

首先在和程序所在路径下创建几个文本文件，我建了两个，文件内容分别为hello python goodbye python 和 i like python。运行程序，得到以下结果

上一篇：python utc datetime转换为时间戳的方法

栏目：Python代码

下一篇：Python中random模块生成随机数详解

本文标题：python实现统计文本中单词出现的频率详解

本文地址：http://www.codeinn.net/misctech/143952.html

更多Python代码

Python代码

python实现统计文本中单词出现的频率详解

阅读排行

推荐教程