Home Backend Development PHP Tutorial Methods and techniques for curl to implement off-site collection_PHP tutorial

Methods and techniques for curl to implement off-site collection_PHP tutorial

Jul 13, 2016 am 10:39 AM
curl

Reason for choosing curl

Regarding curl and file_get_contents, here is an easy-to-understand comparison:
file_get_contents is actually a merged version of a bunch of built-in file operation functions, such as file_exists, fopen, fread, fclose, specially provided for lazy people. And it is mainly used to deal with local files, but because of lazy people, it also adds support for network files;
curl is a library specially used for network interaction, providing a bunch of custom options , used to deal with different environments, and its stability is naturally greater than file_get_contents.

How to use

1. Enable curl support

Since the curl support is not turned on by default after the PHP environment is installed, you need to modify the php.ini file, find; extension=php_curl.dll, remove the colon in front, and restart the service;

2. Use curl to capture data

Copy code The code is as follows:

//Initialize a cURL object
$curl = curl_init();
//Set the URL you need to crawl
curl_setopt($curl, CURLOPT_URL, 'http://www.cmx8.cn');
// Set header
curl_setopt($curl, CURLOPT_HEADER, 1);
//Set cURL parameters and require the results to be saved in a string or output to the screen.
curl_setopt($curl, CURLOPT_RETURNTRANSFER, 1);
// Run cURL and request the web page
$data = curl_exec($curl);
// Close URL request
curl_close($curl );

3. Find key data through regular matching

Copy code The code is as follows:

//$data is the value returned by curl_exec, which is the target content collected
preg_match_all("/
  • (.*?)
  • /",$data, $out, PREG_SET_ORDER);
    foreach($out as $key => $value){
    //Here $value is an array, and records the entire sentence with matching characters and the individually matched characters
    echo 'The entire sentence matched: '.$value[0].'
    ';
    echo 'Single match: '.$value[1].'
    ';
    }

    Tips

    1. Timeout related settings

    You can set some timeout settings through curl_setopt($ch, opt), mainly including:

    CURLOPT_TIMEOUT sets the maximum number of seconds cURL is allowed to execute.
    CURLOPT_TIMEOUT_MS sets the maximum number of milliseconds cURL is allowed to execute. (Added in cURL 7.16.2. Available from PHP 5.2.3.)
    CURLOPT_CONNECTTIMEOUT The time to wait before initiating a connection. If set to 0, it will wait indefinitely.
    CURLOPT_CONNECTTIMEOUT_MS The time to wait for a connection attempt, in milliseconds. If set to 0, wait infinitely. Added in cURL 7.16.2. Available starting with PHP 5.2.3.
    CURLOPT_DNS_CACHE_TIMEOUT sets the time to save DNS information in memory, the default is 120 seconds.

    Copy code The code is as follows:

    curl_setopt($ch, CURLOPT_TIMEOUT, 60); //Only need to set one second The number can be
    curl_setopt($ch, CURLOPT_NOSIGNAL, 1); //Note that the millisecond timeout must be set
    curl_setopt($ch, CURLOPT_TIMEOUT_MS, 200); //The timeout in milliseconds is changed in cURL 7.16.2 join in. Available from PHP 5.2.3

    2. Submit data through post and retain cookies

    Copy code The code is as follows:

    //The following is an example for learning and reference:
    //Curl simulates login discuz program, suitable for DZ7.0

    !extension_loaded('curl') && die( 'The curl extension is not loaded.');

    $discuz_url = 'http://www.lxvoip.com';//Forum address
    $login_url = $discuz_url .'/logging.php ?action=login'; //Login page address
    $get_url = $discuz_url .'/my.php?item=threads'; //My post

    $post_fields = array();
    //The following two items do not need to be modified
    $post_fields['loginfield'] = 'username';
    $post_fields['loginsubmit'] = 'true';
    //Username and password are required Fill in
    $post_fields['username'] = 'lxvoip';
    $post_fields['password'] = '88888888';
    //Security question
    $post_fields['questionid'] = 0 ;
    $post_fields['answer'] = '';
    //@todo verification code
    $post_fields['seccoverify'] = ''; 🎜>$ch = curl_init($login_url);
    curl_setopt($ch, CURLOPT_HEADER, 0);
    curl_setopt($ch, CURLOPT_RETURNTRANSFER, 1); 🎜>curl_close($ch);
    preg_match('//i' , $contents, $matches);
    if(!empty($matches)) {
    $formhash = $matches[1];
    } else {
    die('Not found the forumhash. ');
    }

    //POST data, get COOKIE
    $cookie_file = dirname(__FILE__) . '/cookie.txt';
    //$cookie_file = tempnam('/ tmp');
    $ch = curl_init($login_url);
    curl_setopt($ch, CURLOPT_HEADER, 0);
    curl_setopt($ch, CURLOPT_RETURNTRANSFER, 1); CURLOPT_POST, 1);
    curl_setopt($ch, CURLOPT_POSTFIELDS, $post_fields);
    curl_setopt($ch, CURLOPT_COOKIEJAR, $cookie_file);
    curl_exec($ch);
    curl_close($ch) ;

    //Use the COOKIE obtained above to obtain the content of the page that needs to be logged in to view.
    $ch = curl_init($get_url);
    curl_setopt($ch, CURLOPT_HEADER, 0);
    curl_setopt($ch, CURLOPT_RETURNTRANSFER, 0);
    curl_setopt($ch, CURLOPT_COOKIEFILE, $cookie_file);
    $contents = curl_exec($ch);
    var_dump($contents);






    http://www.bkjia.com/PHPjc/728088.html

    www.bkjia.com
    true

    http: //www.bkjia.com/PHPjc/728088.html

    Reasons for choosing curl Regarding curl and file_get_contents, here is an easy-to-understand comparison: file_get_contents is actually a bunch of built-in Merged versions of file operation functions, such as file_ex...
    Statement of this Website
    The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn

    Hot AI Tools

    Undresser.AI Undress

    Undresser.AI Undress

    AI-powered app for creating realistic nude photos

    AI Clothes Remover

    AI Clothes Remover

    Online AI tool for removing clothes from photos.

    Undress AI Tool

    Undress AI Tool

    Undress images for free

    Clothoff.io

    Clothoff.io

    AI clothes remover

    AI Hentai Generator

    AI Hentai Generator

    Generate AI Hentai for free.

    Hot Article

    R.E.P.O. Energy Crystals Explained and What They Do (Yellow Crystal)
    4 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
    R.E.P.O. Best Graphic Settings
    4 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
    R.E.P.O. How to Fix Audio if You Can't Hear Anyone
    4 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
    WWE 2K25: How To Unlock Everything In MyRise
    1 months ago By 尊渡假赌尊渡假赌尊渡假赌

    Hot Tools

    Notepad++7.3.1

    Notepad++7.3.1

    Easy-to-use and free code editor

    SublimeText3 Chinese version

    SublimeText3 Chinese version

    Chinese version, very easy to use

    Zend Studio 13.0.1

    Zend Studio 13.0.1

    Powerful PHP integrated development environment

    Dreamweaver CS6

    Dreamweaver CS6

    Visual web development tools

    SublimeText3 Mac version

    SublimeText3 Mac version

    God-level code editing software (SublimeText3)

    How to realize the mutual conversion between CURL and python requests in python How to realize the mutual conversion between CURL and python requests in python May 03, 2023 pm 12:49 PM

    Both curl and Pythonrequests are powerful tools for sending HTTP requests. While curl is a command-line tool that allows you to send requests directly from the terminal, Python's requests library provides a more programmatic way to send requests from Python code. The basic syntax for converting curl to Pythonrequestscurl command is as follows: curl[OPTIONS]URL When converting curl command to Python request, we need to convert the options and URL into Python code. Here is an example curlPOST command: curl-XPOST https://example.com/api

    Tutorial on updating curl version under Linux! Tutorial on updating curl version under Linux! Mar 07, 2024 am 08:30 AM

    To update the curl version under Linux, you can follow the steps below: Check the current curl version: First, you need to determine the curl version installed in the current system. Open a terminal and execute the following command: curl --version This command will display the current curl version information. Confirm available curl version: Before updating curl, you need to confirm the latest version available. You can visit curl's official website (curl.haxx.se) or related software sources to find the latest version of curl. Download the curl source code: Using curl or a browser, download the source code file for the curl version of your choice (usually .tar.gz or .tar.bz2

    From start to finish: How to use php extension cURL to make HTTP requests From start to finish: How to use php extension cURL to make HTTP requests Jul 29, 2023 pm 05:07 PM

    From start to finish: How to use php extension cURL for HTTP requests Introduction: In web development, it is often necessary to communicate with third-party APIs or other remote servers. Using cURL to make HTTP requests is a common and powerful way. This article will introduce how to use PHP to extend cURL to perform HTTP requests, and provide some practical code examples. 1. Preparation First, make sure that php has the cURL extension installed. You can execute php-m|grepcurl on the command line to check

    PHP8.1 released: Introducing curl for concurrent processing of multiple requests PHP8.1 released: Introducing curl for concurrent processing of multiple requests Jul 08, 2023 pm 09:13 PM

    PHP8.1 released: Introducing curl for concurrent processing of multiple requests. Recently, PHP officially released the latest version of PHP8.1, which introduced an important feature: curl for concurrent processing of multiple requests. This new feature provides developers with a more efficient and flexible way to handle multiple HTTP requests, greatly improving performance and user experience. In previous versions, handling multiple requests often required creating multiple curl resources and using loops to send and receive data respectively. Although this method can achieve the purpose

    How to handle 301 redirection of web pages in PHP Curl? How to handle 301 redirection of web pages in PHP Curl? Mar 08, 2024 am 11:36 AM

    How to handle 301 redirection of web pages in PHPCurl? When using PHPCurl to send network requests, you will often encounter a 301 status code returned by the web page, indicating that the page has been permanently redirected. In order to handle this situation correctly, we need to add some specific options and processing logic to the Curl request. The following will introduce in detail how to handle 301 redirection of web pages in PHPCurl, and provide specific code examples. 301 redirect processing principle 301 redirect means that the server returns a 30

    what is linux curl what is linux curl Apr 20, 2023 pm 05:05 PM

    In Linux, curl is a very practical tool for transferring data to and from the server. It is a file transfer tool that uses URL rules to work under the command line; it supports file upload and download, and is a comprehensive transfer tool. . Curl provides a lot of very useful functions, including proxy access, user authentication, ftp upload and download, HTTP POST, SSL connection, cookie support, breakpoint resume and so on.

    How to set cookies in php curl How to set cookies in php curl Sep 26, 2021 am 09:27 AM

    How to set cookies in php curl: 1. Create a PHP sample file; 2. Set cURL transmission options through the "curl_setopt" function; 3. Pass the cookie in CURL.

    Solution to PHP Fatal error: Call to undefined function curl_setopt() Solution to PHP Fatal error: Call to undefined function curl_setopt() Jun 23, 2023 am 08:18 AM

    PHP is a widely used open source scripting language used by many websites. However, sometimes you may encounter the problem PHPFatalerror:Calltoundefinedfunctioncurl_setopt(), which may prevent your website from working properly. So what exactly causes this problem? In PHP, curl_setopt() is a very important function, which is used to extend the library through curl

    See all articles