Home >php教程 >php手册 >采集论坛程序:模拟登陆,抓取页面

采集论坛程序:模拟登陆,抓取页面

WBOY
WBOYOriginal
2016-06-13 10:36:45808browse

//    吴燕军
//   2009-06-27
//   采集程序php
set_time_limit(0);

//cookie保存目录
$cookie_jar = /tmp/cookie.tmp;

/*函数------------------------------------------------------------------------------------------------------------*/

//模拟请求数据
function request($url,$postfields,$cookie_jar,$referer){
$ch = curl_init();
$options = array(CURLOPT_URL => $url,
      CURLOPT_HEADER => 0,
      CURLOPT_NOBODY => 0,
      CURLOPT_PORT => 80,
      CURLOPT_POST => 1,
      CURLOPT_POSTFIELDS => $postfields,
      CURLOPT_RETURNTRANSFER => 1,
      CURLOPT_FOLLOWLOCATION => 1,
      CURLOPT_COOKIEJAR => $cookie_jar,
      CURLOPT_COOKIEFILE => $cookie_jar,
      CURLOPT_REFERER => $referer
);
curl_setopt_array($ch, $options);
$code = curl_exec($ch);
curl_close($ch);
return $code;
}

//获取帖子列表
function getThreadsList($code){
preg_match_all(/

Statement:
The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn