|
|
|
|
|
獲得網(wǎng)頁header信息,是網(wǎng)站開發(fā)人員和維護人員常用的技術(shù)。網(wǎng)頁的header信息,非常豐富,非專業(yè)人士一般較難讀懂和理解各個項目的含義。
獲取網(wǎng)頁header信息,方法多種多樣,就php語言來說,我作為一個菜鳥,知道的方法也有4種那么多。下面逐一獻上。
方法一:使用get_headers()函數(shù)
這個方法很多人使用,也很簡單便捷,只需要兩行代碼即可搞定。如下:
$thisurl = "http://gazebo2go.com/";
print_r(get_headers($thisurl, 1));
得到的結(jié)果為:
Array
(
[0] => HTTP/1.1 200 OK
[Cache-Control] => max-age=86400
[Content-Length] => 76102
[Content-Type] => text/html
[Content-Location] => http://gazebo2go.com/index.html
[Last-Modified] => Fri, 19 Jul 2013 03:52:30 GMT
[Accept-Ranges] => bytes
[ETag] => "50bc48643384ce1:5cb3"
[Server] => Microsoft-IIS/6.0
[X-Powered-By] => ASP.NET
[Date] => Fri, 19 Jul 2013 09:06:39 GMT
[Connection] => close
)
方法二:使用http_response_header
代碼也很簡單,僅需三行:
$thisurl = "http://gazebo2go.com/";
$html = file_get_contents($thisurl );
print_r($http_response_header);
得到的結(jié)果為:
Array
(
[0] => HTTP/1.1 200 OK
[1] => Cache-Control: max-age=86400
[2] => Content-Length: 76102
[3] => Content-Type: text/html
[4] => Content-Location: http://gazebo2go.com/index.html
[5] => Last-Modified: Fri, 19 Jul 2013 03:52:30 GMT
[6] => Accept-Ranges: bytes
[7] => ETag: "50bc48643384ce1:5cb3"
[8] => Server: Microsoft-IIS/6.0
[9] => X-Powered-By: ASP.NET
[10] => Date: Fri, 19 Jul 2013 09:06:41 GMT
[11] => Connection: close
)
方法三:使用stream_get_meta_data()函數(shù)
代碼也只有三行:
$thisurl = "http://gazebo2go.com/";
$fp = fopen($thisurl, 'r');
print_r(stream_get_meta_data($fp));
得到的結(jié)果為:
Array
(
[wrapper_data] => Array
(
[0] => HTTP/1.1 200 OK
[1] => Cache-Control: max-age=86400
[2] => Content-Length: 76102
[3] => Content-Type: text/html
[4] => Content-Location: http://gazebo2go.com/index.html
[5] => Last-Modified: Fri, 19 Jul 2013 03:52:30 GMT
[6] => Accept-Ranges: bytes
[7] => ETag: "50bc48643384ce1:5cb3"
[8] => Server: Microsoft-IIS/6.0
[9] => X-Powered-By: ASP.NET
[10] => Date: Fri, 19 Jul 2013 09:06:41 GMT
[11] => Connection: close
)
[wrapper_type] => http
[stream_type] => tcp_socket
[mode] => r+
[unread_bytes] => 1086
[seekable] =>
[uri] => http://gazebo2go.com/
[timed_out] =>
[blocked] => 1
[eof] =>
)
上述三種方法都可以輕松獲得網(wǎng)頁header信息,且包含的信息都已經(jīng)相當豐富,滿足一般要求,不過比較遺憾的是,上述三種方法都不能用來檢測網(wǎng)頁是否啟用了GZip壓縮。要檢測GZip壓縮,還需其他的方法才行。這里介紹的是用curl()函數(shù)來檢測。
使用curl獲得header可以檢測GZip壓縮
先貼出代碼:
<?php
$szUrl = 'http://gazebo2go.com/';
$curl = curl_init();
curl_setopt($curl, CURLOPT_URL, $szUrl);
curl_setopt($curl, CURLOPT_HEADER, 1); //輸出header信息
curl_setopt($curl, CURLOPT_RETURNTRANSFER, 1); //不顯示網(wǎng)頁內(nèi)容
curl_setopt($curl, CURLOPT_ENCODING, ''); //允許執(zhí)行g(shù)zip
$data=curl_exec($curl);
if(!curl_errno($curl))
{
$info = curl_getinfo($curl);
$httpHeaderSize = $info['header_size']; //header字符串體積
$pHeader = substr($data, 0, $httpHeaderSize); //獲得header字符串
$split = array("\r\n", "\n", "\r"); //需要格式化header字符串
$pHeader = str_replace($split, '<br>', $pHeader); //使用<br>換行符格式化輸出到網(wǎng)頁上
echo $pHeader;
}
?>
輸出結(jié)果如下:
HTTP/1.1 200 OK
Cache-Control: max-age=86400
Content-Length: 15189
Content-Type: text/html
Content-Encoding: gzip
Content-Location: http://gazebo2go.com/index.html
Last-Modified: Fri, 19 Jul 2013 03:52:28 GMT
Accept-Ranges: bytes
ETag: "0268633384ce1:5cb3"
Vary: Accept-Encoding
Server: Microsoft-IIS/6.0
X-Powered-By: ASP.NET
Date: Fri, 19 Jul 2013 09:27:21 GMT
上面輸出結(jié)果里可以看到一個項目:Content-Encoding: gzip,這個正是我們用來判斷網(wǎng)頁是否啟用GZip壓縮的項目。
另外,需要認真注意下本實例里的注釋部分,不能少了任何一項,否則可能獲取header信息有誤。