Home Backend Development PHP Tutorial PHP implementation code for reading large files_PHP tutorial

PHP implementation code for reading large files_PHP tutorial

Jul 21, 2016 pm 03:13 PM
php code exist accomplish Quick operate document Way hour most of read conduct

In PHP, when reading files, the fastest way is to use some functions such as file and file_get_contents. A few simple lines of code can beautifully complete the functions we need. But when the file being operated is a relatively large file, these functions may be insufficient. The following will start with a requirement to explain the commonly used operating methods when reading large files.
Requirements

There is an 800M log file with about 5 million lines. Use PHP to return the contents of the last few lines.

Implementation method

1. Directly use the file function to operate

Note: Since the file function reads all the contents into the memory at once, and In order to prevent some poorly written programs from taking up too much memory and causing insufficient system memory and causing the server to crash, PHP limits the maximum memory usage to 16M by default. This is done through memory_limit = 16M in php.ini. To set, if this value is set to -1, the memory usage is not limited.

The following is a piece of code that uses file to extract the last line of this file.
The entire code execution takes 116.9613 (s).

Copy code The code is as follows:

$fp = fopen($file, "r");
$num = 10;
$chunk = 4096;
$fs = sprintf("%u", filesize($file));
$max = (intval($fs) == PHP_INT_MAX ) ? PHP_INT_MAX : filesize($file);
for ($len = 0; $len < $max; $len += $chunk) {
$seekSize = ($max - $len > $ chunk) ? $chunk : $max - $len;
fseek($fp, ($len + $seekSize) * -1, SEEK_END);
$readData = fread($fp, $seekSize) . $ readData;

if (substr_count($readData, "n") >= $num + 1) {
preg_match("!(.*?n){".($num)." }$!", $readData, $match);
$data = $match[0];
break;
}
}
fclose($fp);
echo $data;

My machine has 2G of memory. When I press F5 to run, the system turns gray and recovers after almost 20 minutes. It can be seen that all such large files can be read directly. into the memory, the consequences are quite serious, so it is not a last resort. The memory_limit cannot be adjusted too high, otherwise the only choice is to call the computer room and reset the machine.

2. Directly call Linux The tail command displays the last few lines of the log file

Under the Linux command line, you can directly use tail -n 10 access.log to easily display the last few lines of the log file. You can directly use php to call the tail command. , execute the php code as follows.
The entire code execution takes 0.0034 (s)
Copy the code The code is as follows:

file = 'access.log';
$file = escapeshellarg($file); // Safely escape command line parameters
$line = `tail -n 1 $file`;
echo $line;


3. Directly use php’s fseek to perform file operations

This method is the most common method, it does not require Read all the contents of the file into the content, but operate directly through pointers, so the efficiency is quite efficient. When using fseek to operate the file, there are many different methods, and the efficiency may be slightly different, as follows There are two commonly used methods.

Method 1
First find the last EOF of the file through fseek, then find the starting position of the last line, and get the data of this line. Find the starting position of the next row, then take the position of this row, and so on, until the $num row is found.
The implementation code is as follows
The entire code execution takes 0.0095 (s)
Copy the code The code is as follows:

function tail($fp,$n,$base=5)
{
assert($n>0);
$pos = $n+1;
$lines = array() ;
while(count($lines)< =$n){
try{
fseek($fp,-$pos,SEEK_END);
} catch (Exception $e){
fseek(0);
break;
}
$pos *= $base;
while(!feof($fp)){
array_unshift($lines,fgets($ fp));
}
}
return array_slice($lines,0,$n);
}
var_dump(tail(fopen("access.log","r+") ,10));

Method 2
Still use fseek to read from the end of the file, but this time it is not read one by one, but one block Read one block, each time a block of data is read, the read data is placed in a buf, and then the number of newline characters (n) is used to determine whether the last $num rows of data have been read.
Implementation code As follows
The entire code execution takes 0.0009(s).
Copy the code The code is as follows:

$fp = fopen($file, "r");
$line = 10;
$pos = -2;
$t = " ";
$data = "";
while ( $line > 0) {
while ($t != "n") {
fseek($fp, $pos, SEEK_END);
$t = fgetc($fp);
$pos --;
}
$t = " ";
$data .= fgets($fp);
$line --;
}
fclose ($fp );
echo $data

Method Three
The entire code execution takes 0.0003(s)
Copy the code The code is as follows:

ini_set('memory_limit','-1');
$file = 'access.log';
$data = file($file);
$line = $data [count($data)-1];
echo $line;

www.bkjia.comtruehttp: //www.bkjia.com/PHPjc/326534.htmlTechArticleIn php, when reading files, the fastest way is to use something like file, file_get_contents Class functions, just a few lines of code can beautifully complete our...
Statement of this Website
The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

AI Hentai Generator

AI Hentai Generator

Generate AI Hentai for free.

Hot Article

R.E.P.O. Energy Crystals Explained and What They Do (Yellow Crystal)
2 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
R.E.P.O. Best Graphic Settings
2 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
R.E.P.O. How to Fix Audio if You Can't Hear Anyone
2 weeks ago By 尊渡假赌尊渡假赌尊渡假赌

Hot Tools

Notepad++7.3.1

Notepad++7.3.1

Easy-to-use and free code editor

SublimeText3 Chinese version

SublimeText3 Chinese version

Chinese version, very easy to use

Zend Studio 13.0.1

Zend Studio 13.0.1

Powerful PHP integrated development environment

Dreamweaver CS6

Dreamweaver CS6

Visual web development tools

SublimeText3 Mac version

SublimeText3 Mac version

God-level code editing software (SublimeText3)

CakePHP Project Configuration CakePHP Project Configuration Sep 10, 2024 pm 05:25 PM

In this chapter, we will understand the Environment Variables, General Configuration, Database Configuration and Email Configuration in CakePHP.

PHP 8.4 Installation and Upgrade guide for Ubuntu and Debian PHP 8.4 Installation and Upgrade guide for Ubuntu and Debian Dec 24, 2024 pm 04:42 PM

PHP 8.4 brings several new features, security improvements, and performance improvements with healthy amounts of feature deprecations and removals. This guide explains how to install PHP 8.4 or upgrade to PHP 8.4 on Ubuntu, Debian, or their derivati

CakePHP Date and Time CakePHP Date and Time Sep 10, 2024 pm 05:27 PM

To work with date and time in cakephp4, we are going to make use of the available FrozenTime class.

CakePHP File upload CakePHP File upload Sep 10, 2024 pm 05:27 PM

To work on file upload we are going to use the form helper. Here, is an example for file upload.

CakePHP Routing CakePHP Routing Sep 10, 2024 pm 05:25 PM

In this chapter, we are going to learn the following topics related to routing ?

Discuss CakePHP Discuss CakePHP Sep 10, 2024 pm 05:28 PM

CakePHP is an open-source framework for PHP. It is intended to make developing, deploying and maintaining applications much easier. CakePHP is based on a MVC-like architecture that is both powerful and easy to grasp. Models, Views, and Controllers gu

CakePHP Creating Validators CakePHP Creating Validators Sep 10, 2024 pm 05:26 PM

Validator can be created by adding the following two lines in the controller.

CakePHP Working with Database CakePHP Working with Database Sep 10, 2024 pm 05:25 PM

Working with database in CakePHP is very easy. We will understand the CRUD (Create, Read, Update, Delete) operations in this chapter.

See all articles