WAP网站是站啥啥
?wap是挪动端仍是手机端?python百度下拉框关头词华集源码:
07 | url=f"https://www.百度.com/sugrec?pre=1&ie=utf-8&json=1&prod=pc&wd={ word}" |
08 | html=requests.get(url) |
13 | forkey_word inhtml['g']: |
15 | key_words.append(key_word['q']) |
20 | url ='https://sp0.百度.com/5a1Fazu8AA54nxGko9WTAnF6hhy/su?wd=%s&sugmode=2&json=1&p=3&sid=1427_21091_21673_22581&req=2&pbs=%%E5%%BF%%AB%%E6%%89%%8B&csor=2&pwd=%%E5%%BF%%AB%%E6%%89%%8B&cb=jQuery11020924966752020363_1498055470768&_=1498055470781'%word |
21 | r =requests.get(url, verify=False) |
23 | res =cont[41: -2].decode('gbk') |
24 | res_json =json.loads(res) |
28 | url=f'http://suggestion.百度.com/su?仍手wd={ word}&sugmode=3&json=1' |
29 | html=requests.get(url).text |
30 | html=html.replace("window.百度.sug(",'') |
31 | html =html.replace(")", '') |
32 | html =html.replace(";", '') |
34 | html =json.loads(html) |
40 | opencsv=open('word.csv','a+') |
41 | forword inopen('gjc.txt',encoding='utf-8'): |
42 | print(urllib.parse.quote_plus(word)) |
43 | url='https://sp0.百度.com/5a1Fazu8AA54nxGko9WTAnF6hhy/su?wd=%s&sugmode=2&json=1&p=3&sid=1427_21091_21673_22581&req=2'%urllib.parse.quote_plus(word) |
44 | html=requests.get(url).text |
45 | html=html.replace('window.百度.sug(','') |
46 | html=html.replace(');','') |
52 | opencsv.write('%s\n'%i) |
54 | defget_more_word(word): |
56 | fori in'abcdefghijklmnopqrstuvwxyz': |
57 | more_word.extend(get_keywords('%s%s'%(word,i))) |
60 | print(len(list(set(more_word)))) |
61 | returnlist(set(more_word)) |
66 | fori in'abcdefghijklmnopqrstuvwxyz': |
67 | all_words +=get_sug(word+i) |
68 | print(len(list(set(all_words)))) |
69 | returnlist(set(all_words)) |
|
供给多种python百度下拉框关头词华团编制,基于百度API接口完成 ,机端可导出到Excel表格
,站啥本文供给4个群集函数及两个汇总函数,动端屈就本身的仍手需求矫捷独霸。
机端