Python3使用requests套件抓取並保存網頁原始碼的方法介紹-Python教學-PHP中文網

首頁

後端開發

Python教學

Python3使用requests套件抓取並保存網頁原始碼的方法介紹

高洛峰

Mar 07, 2017 pm 03:50 PM

本文實例講述了Python3使用requests套件抓取並保存網頁原始碼的方法。分享給大家供大家參考，具體如下：

使用Python 3的requests模組抓取網頁原始碼並儲存到檔案範例：

import requests
html = requests.get("http://www.baidu.com")
with open(&#39;test.txt&#39;,&#39;w&#39;,encoding=&#39;utf-8&#39;) as f:
 f.write(html.text)

#這是一個基本的檔案保存操作，但這裡有幾個值得注意的問題：

1.安裝requests包，命令列輸入pip install requests即可自動安裝。很多人推薦使用requests，自帶的urllib.request也可以抓取網頁原始碼

2.open方法encoding參數設為utf-8，否則儲存的檔案會出現亂碼。

3.如果直接在cmd中輸出抓取的內容，會提示各種編碼錯誤，所以儲存到檔案檢視。

4.with open方法是更好的寫法，可以自動操作完畢後釋放資源。

另一個例子：

import requests
ff = open(&#39;testt.txt&#39;,&#39;w&#39;,encoding=&#39;utf-8&#39;)
with open(&#39;test.txt&#39;,encoding="utf-8") as f:
 for line in f:
 ff.write(line)
ff.close()

這是一個演示讀取一個txt文件，每次讀取一行，並保存到另一個txt文件中的範例。

因為在命令列中列印每次讀取一行的數據，中文會出現編碼錯誤，所以每次讀取一行並保存到另一個文件，這樣來測試讀取是否正常。（注意open的時候制定encoding編碼方式）

更多Python3使用requests包抓取並保存網頁源碼的方法介紹相關文章請關注PHP中文網！

陳述

本文內容由網友自願投稿，版權歸原作者所有。本站不承擔相應的法律責任。如發現涉嫌抄襲或侵權的內容，請聯絡admin@php.cn

Python的混合方法：編譯和解釋合併May 08, 2025 am 12:16 AM

pythonuseshybridapprace，ComminingCompilationTobyTecoDeAndInterpretation.1）codeiscompiledtoplatform-Indepententbybytecode.2）bytecodeisisterpretedbybythepbybythepythonvirtualmachine，增強效率和通用性。

了解python的' for”和' then”循環之間的差異May 08, 2025 am 12:11 AM

theKeyDifferencesBetnewpython's“ for”和“ for”和“ loopsare：1）” for“ loopsareIdealForiteringSequenceSquencesSorkNowniterations，而2）”，而“ loopsareBetterforConterContinuingUntilacTientInditionIntionismetismetistismetistwithOutpredefinedInedIterations.un

Python串聯列表與重複May 08, 2025 am 12:09 AM

在Python中，可以通過多種方法連接列表並管理重複元素：1)使用運算符或extend()方法可以保留所有重複元素；2)轉換為集合再轉回列表可以去除所有重複元素，但會丟失原有順序；3)使用循環或列表推導式結合集合可以去除重複元素並保持原有順序。

Python列表串聯性能：速度比較May 08, 2025 am 12:09 AM

fasteStmethodMethodMethodConcatenationInpythondependersonListsize：1）forsmalllists，operatorseffited.2）forlargerlists，list.extend.extend（）orlistComprechensionfaster，withextendEffaster，withExtendEffers，withextend（）withextend（）是extextend（）asmoremory-ememory-emmoremory-emmoremory-emmodifyinginglistsin-place-place-place。

您如何將元素插入python列表中？May 08, 2025 am 12:07 AM

toInSerteLementIntoApythonList，useAppend（）toaddtotheend，insert（）foreSpificPosition，andextend（）formultiplelements.1）useappend（）foraddingsingleitemstotheend.2）useAddingsingLeitemStotheend.2）useeapecificindex，toadapecificindex，toadaSpecificIndex，toadaSpecificIndex，blyit'ssssssslorist.3 toaddextext.3

Python是否列表動態陣列或引擎蓋下的鏈接列表？May 07, 2025 am 12:16 AM

pythonlistsareimplementedasdynamicarrays，notlinkedlists.1）他們areStoredIncoNtiguulMemoryBlocks，mayrequireRealLealLocationWhenAppendingItems，EmpactingPerformance.2）LinkesedlistSwoldOfferefeRefeRefeRefeRefficeInsertions/DeletionsButslowerIndexeDexedAccess，Lestpypytypypytypypytypy

如何從python列表中刪除元素？May 07, 2025 am 12:15 AM

pythonoffersFourmainMethodStoreMoveElement Fromalist：1）刪除（值）emovesthefirstoccurrenceofavalue，2）pop（index）emovesanderturnsanelementataSpecifiedIndex，3）delstatementremoveselemsbybybyselementbybyindexorslicebybyindexorslice，and 4）

試圖運行腳本時，應該檢查是否會遇到'權限拒絕”錯誤？May 07, 2025 am 12:12 AM

toresolvea“ dermissionded”錯誤Whenrunningascript，跟隨台詞：1）CheckAndAdjustTheScript'Spermissions ofchmod xmyscript.shtomakeitexecutable.2）nesureThEseRethEserethescriptistriptocriptibationalocatiforecationAdirectorywherewhereyOuhaveWritePerMissionsyOuhaveWritePermissionsyYouHaveWritePermissions，susteSyAsyOURHomeRecretectory。

See all articles