亮点: 1、利用php也能实现对页面div的切割处理。这里的做法抛砖引玉,希望读者能够提供更加完美的解决方案。 2、切割处理方法已经封装成一个方法,可以直接引用。 3、顺便加上的截
亮点:
1、利用php也能实现对页面div的切割处理。这里的做法抛砖引玉,希望读者能够提供更加完美的解决方案。
2、切割处理方法已经封装成一个方法,可以直接引用。
3、顺便加上的截取。//getWebDiv('id="taglist"','http://www.cnblogs.com/Zjmainstay/tag/');
View Code
<span>php </span><span>header</span>("Content-type: text/html; charset=utf-8"<span>); </span><span>function</span> getWebDiv(<span>$div_id</span>,<span>$url</span>=<span>false</span>,<span>$data</span>=<span>false</span><span>){ </span><span>if</span>(<span>$url</span> !== <span>false</span><span>){ </span><span>$data</span> = <span>file_get_contents</span>( <span>$url</span><span> ); } </span><span>$charset_pos</span> = <span>stripos</span>(<span>$data</span>,'charset'<span>); </span><span>if</span>(<span>$charset_pos</span><span>) { </span><span>if</span>(<span>stripos</span>(<span>$data</span>,'charset=utf-8',<span>$charset_pos</span><span>)) { </span><span>$data</span> = <span>iconv</span>('utf-8','utf-8',<span>$data</span><span>); }</span><span>else</span> <span>if</span>(<span>stripos</span>(<span>$data</span>,'charset=gb2312',<span>$charset_pos</span><span>)) { </span><span>$data</span> = <span>iconv</span>('gb2312','utf-8',<span>$data</span><span>); }</span><span>else</span> <span>if</span>(<span>stripos</span>(<span>$data</span>,'charset=gbk',<span>$charset_pos</span><span>)) { </span><span>$data</span> = <span>iconv</span>('gbk','utf-8',<span>$data</span><span>); } } </span><span>preg_match_all</span>('/<div>$data,<span>$pre_matches</span>,PREG_OFFSET_CAPTURE); <span>//</span><span>获取所有div前缀</span> <span>preg_match_all</span>('/$data,<span>$suf_matches</span>,PREG_OFFSET_CAPTURE); <span>//</span><span>获取所有div后缀</span> <span>$hit</span> = <span>strpos</span>(<span>$data</span>,<span>$div_id</span><span>); </span><span>if</span>(<span>$hit</span> == -1) <span>return</span> <span>false</span>; <span>//</span><span>未命中</span> <span>$divs</span> = <span>array</span>(); <span>//</span><span>合并所有div</span> <span>foreach</span>(<span>$pre_matches</span>[0] <span>as</span> <span>$index</span>=><span>$pre_div</span><span>){ </span><span>$divs</span>[(int)<span>$pre_div</span>[1]] = 'p'<span>; </span><span>$divs</span>[(int)<span>$suf_matches</span>[0][<span>$index</span>][1]] = 's'<span>; } </span><span>//</span><span>对div进行排序</span> <span>$sort</span> = <span>array_keys</span>(<span>$divs</span><span>); </span><span>asort</span>(<span>$sort</span><span>); </span><span>$count</span> = <span>count</span>(<span>$pre_matches</span>[0<span>]); </span><span>foreach</span>(<span>$pre_matches</span>[0] <span>as</span> <span>$index</span>=><span>$pre_div</span><span>){ </span><span>//</span><span><div> <span>if</span>((<span>$pre_matches</span>[0][<span>$index</span>][1] $hit) && (<span>$hit</span> $pre_matches[0][<span>$index</span>+1][1<span>])){ </span><span>$deeper</span> = 0<span>; </span><span>//</span><span>弹出被命中div前的div</span> <span>while</span>(<span>array_shift</span>(<span>$sort</span>) != <span>$pre_matches</span>[0][<span>$index</span>][1] && (<span>$count</span>--)) <span>continue</span><span>; </span><span>//</span><span>对剩余div进行匹配,若下一个为前缀,则向下一层,$deeper加1, //否则后退一层,$deeper减1,$deeper为0则命中匹配,计算div长度</span> <span>foreach</span>(<span>$sort</span> <span>as</span> <span>$key</span><span>){ </span><span>if</span>(<span>$divs</span>[<span>$key</span>] == 'p') <span>$deeper</span>++<span>; </span><span>else</span> <span>if</span>(<span>$deeper</span> == 0<span>) { </span><span>$length</span> = <span>$key</span>-<span>$pre_matches</span>[0][<span>$index</span>][1<span>]; </span><span>break</span><span>; }</span><span>else</span><span> { </span><span>$deeper</span>--<span>; } } </span><span>$hitDivString</span> = <span>substr</span>(<span>$data</span>,<span>$pre_matches</span>[0][<span>$index</span>][1],<span>$length</span>).'</div>'<span>; </span><span>break</span><span>; } } </span><span>return</span> <span>$hitDivString</span><span>; } </span><span>echo</span> getWebDiv('id="taglist"','http://www.cnblogs.com/Zjmainstay/tag/'<span>); </span><span>//</span><span>End_php</span> <p>考虑到id符号问题,id="u"由用户自己填写。</p> <p>声明:此段php只针对带 id div内容的读取。</p> <p><img src="/static/imghwm/default1.png" data-src="/inc/test.jsp?url=http%3A%2F%2Fimages.cnblogs.com%2FOutliningIndicators%2FContractedBlock.gif&refer=http%3A%2F%2Fwww.cnblogs.com%2FZjmainstay%2Farchive%2F2012%2F08%2F06%2Fphp_getDivContain.html" class="lazy" alt="php获取页面并切割页面div内容" ><img src="/static/imghwm/default1.png" data-src="/inc/test.jsp?url=http%3A%2F%2Fimages.cnblogs.com%2FOutliningIndicators%2FExpandedBlockStart.gif&refer=http%3A%2F%2Fwww.cnblogs.com%2FZjmainstay%2Farchive%2F2012%2F08%2F06%2Fphp_getDivContain.html" class="lazy" alt="php获取页面并切割页面div内容" ><span>View Code </span> </p> <p> </p> <pre class="brush:php;toolbar:false"><span> 1</span> <span>php </span><span> 2</span> <span>header</span>("Content-type: text/html; charset=utf-8"<span>); </span><span> 3</span> <span>function</span> getWebTag(<span>$tag_id</span>,<span>$url</span>=<span>false</span>,<span>$tag</span>='div',<span>$data</span>=<span>false</span><span>){ </span><span> 4</span> <span>if</span>(<span>$url</span> !== <span>false</span><span>){ </span><span> 5</span> <span>$data</span> = <span>file_get_contents</span>( <span>$url</span><span> ); </span><span> 6</span> <span> } </span><span> 7</span> <span>$charset_pos</span> = <span>stripos</span>(<span>$data</span>,'charset'<span>); </span><span> 8</span> <span>if</span>(<span>$charset_pos</span><span>) { </span><span> 9</span> <span>if</span>(<span>stripos</span>(<span>$data</span>,'charset=utf-8',<span>$charset_pos</span><span>)) { </span><span>10</span> <span>$data</span> = <span>iconv</span>('utf-8','utf-8',<span>$data</span><span>); </span><span>11</span> }<span>else</span> <span>if</span>(<span>stripos</span>(<span>$data</span>,'charset=gb2312',<span>$charset_pos</span><span>)) { </span><span>12</span> <span>$data</span> = <span>iconv</span>('gb2312','utf-8',<span>$data</span><span>); </span><span>13</span> }<span>else</span> <span>if</span>(<span>stripos</span>(<span>$data</span>,'charset=gbk',<span>$charset_pos</span><span>)) { </span><span>14</span> <span>$data</span> = <span>iconv</span>('gbk','utf-8',<span>$data</span><span>); </span><span>15</span> <span> } </span><span>16</span> <span> } </span><span>17</span> <span>18</span> <span>preg_match_all</span>('/$tag.'/i',$data,$pre_matches,PREG_OFFSET_CAPTURE); //获取所有div前缀 19 preg_match_all('/$tag.'/i',$data,$suf_matches,PREG_OFFSET_CAPTURE); //获取所有div后缀 20 $hit = strpos($data,$tag_id); 21 if($hit == -1) return false; //未命中 22 $divs = array(); //合并所有div 23 foreach($pre_matches[0] as $index=>$pre_div){ 24 $divs[(int)$pre_div[1]] = 'p'; 25 $divs[(int)$suf_matches[0][$index][1]] = 's'; 26 } 27 28 //对div进行排序 29 $sort = array_keys($divs); 30 asort($sort); 31 32 $count = count($pre_matches[0]); 33 foreach($pre_matches[0] as $index=>$pre_div){ 34 //
修复:stripos($data,'charset=utf-8',$charset_pos) 加入charset=,避免有些gb2312格式的网页中包含utf-8造成错误。或者用户可以自行修改函数传入一个确定的charset参数。
演示地址:parseDiv

php把负数转为正整数的方法:1、使用abs()函数将负数转为正数,使用intval()函数对正数取整,转为正整数,语法“intval(abs($number))”;2、利用“~”位运算符将负数取反加一,语法“~$number + 1”。

实现方法:1、使用“sleep(延迟秒数)”语句,可延迟执行函数若干秒;2、使用“time_nanosleep(延迟秒数,延迟纳秒数)”语句,可延迟执行函数若干秒和纳秒;3、使用“time_sleep_until(time()+7)”语句。

php字符串有下标。在PHP中,下标不仅可以应用于数组和对象,还可应用于字符串,利用字符串的下标和中括号“[]”可以访问指定索引位置的字符,并对该字符进行读写,语法“字符串名[下标值]”;字符串的下标值(索引值)只能是整数类型,起始值为0。

php除以100保留两位小数的方法:1、利用“/”运算符进行除法运算,语法“数值 / 100”;2、使用“number_format(除法结果, 2)”或“sprintf("%.2f",除法结果)”语句进行四舍五入的处理值,并保留两位小数。

判断方法:1、使用“strtotime("年-月-日")”语句将给定的年月日转换为时间戳格式;2、用“date("z",时间戳)+1”语句计算指定时间戳是一年的第几天。date()返回的天数是从0开始计算的,因此真实天数需要在此基础上加1。

在php中,可以使用substr()函数来读取字符串后几个字符,只需要将该函数的第二个参数设置为负值,第三个参数省略即可;语法为“substr(字符串,-n)”,表示读取从字符串结尾处向前数第n个字符开始,直到字符串结尾的全部字符。

方法:1、用“str_replace(" ","其他字符",$str)”语句,可将nbsp符替换为其他字符;2、用“preg_replace("/(\s|\ \;||\xc2\xa0)/","其他字符",$str)”语句。

查找方法:1、用strpos(),语法“strpos("字符串值","查找子串")+1”;2、用stripos(),语法“strpos("字符串值","查找子串")+1”。因为字符串是从0开始计数的,因此两个函数获取的位置需要进行加1处理。


Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

AI Hentai Generator
Generate AI Hentai for free.

Hot Article

Hot Tools

SublimeText3 Chinese version
Chinese version, very easy to use

SAP NetWeaver Server Adapter for Eclipse
Integrate Eclipse with SAP NetWeaver application server.

VSCode Windows 64-bit Download
A free and powerful IDE editor launched by Microsoft

Dreamweaver CS6
Visual web development tools

SublimeText3 Mac version
God-level code editing software (SublimeText3)