Word histogram

时间：2014-08-17 15:30:32 阅读：338 评论：0 收藏：0 [点我收藏+]

Here is a program that reads a file and builds a histogram of the words in the file:

bubuko.com,布布扣

process_file loops through the lines of the file, passing them one at a time to process_line. The histogram h is being used as an accumulator. process_line uses the string method replace to replace hyphens with spaces before using split to break the line into a list of strings. It traverses the list of words and uses strip and lower to remove punctuation and convert to lower case. (It is a shorthand to say that strings are ‘converted;’ remember that string are immutable, so methods like strip and lower return new strings.)

Finally, process_line updates the histogram by creating a new item incrementing an existing one. To count the total number of words in the file, we can add up the frequencies in the histogram:

bubuko.com,布布扣

from Thinking in Python

Word histogram,布布扣,bubuko.com

Word histogram

原文：http://www.cnblogs.com/ryansunyu/p/3917789.html

踩

(0)

评论一句话评论（0）

分享档案

更多>

2021年09月23日 (328)
2021年09月24日 (313)
2021年09月17日 (191)
2021年09月15日 (369)
2021年09月16日 (411)
2021年09月13日 (439)
2021年09月11日 (398)
2021年09月12日 (393)
2021年09月10日 (160)
2021年09月08日 (222)