


Using Symfony's Crawler component to analyze HTML_php instances in laravel
This article mainly introduces the use of Symfony's Crawler component to analyze HTML in laravel. Friends in need can refer to it
The full name of Crawler is DomCrawler, which is a component of the Symfony framework. What is outrageous is that DomCrawler does not have Chinese documentation, and Symfony has not translated this part, so development using DomCrawler can only be explored bit by bit. Now I will summarize the experience in the use process.
The first thing is to install
composer require symfony/dom-crawler composer require symfony/css-selector
css-seelctor is the css selector, some functions will be used when selecting nodes with css
The example used in the manual is
use Symfony\Component\DomCrawler\Crawler; $html = <<<‘HTML‘ Hello World! Hello Crawler! HTML; $crawler = new Crawler($html); foreach ($crawler as $domElement) { var_dump($domElement->nodeName); }
The printed result is
string ‘html‘ (length=4)
Because the nodeName of this html code is html, and my English is not good, I thought the program was wrong when I started using it. . .
In the actual use process, if new Crawler ($html) will have garbled code problem, it should be related to the page encoding, so you can use the following method, first initialize the crawler, and then add node
$crawler = new Crawler(); $crawler->addHtmlContent($html);
The second parameter of addHtmlContent is charset, and the default is utf-8.
For other examples, please refer to the official documentation, http://symfony.com/doc/current/components/dom_crawler.html
Record the work and try it out bit by bit. Usage
filterXPath(string $xpath) method, according to the manual, the parameter of this method is $xpath, and p, p and other blocks are often used.
echo $crawler->filterXPath(‘//body/p‘)->text(); echo $crawler->filterXPath(‘//body/p‘)->last()->text();
The output is the text of the first and next p tag block
var_dump($crawler->filterXPath(‘//body‘)->html());
Output the html in the body
foreach ($crawler->filterXPath(‘//body/p‘) as $i => $node) { $c = new Crawler($node); echo $c->filter(‘p‘)->text(); }
filterXPath obtains an array of DOMElement blocks, each The DOMElement block can use the new crawler object to continue parsing
$nodeValues = $crawler->filterXPath(‘//body/p‘)->each(function (Crawler $node, $i) { return $node->text(); });
crawler provides an each loop and uses closure functions to simplify the code. However, note that this way of writing $nodeValues gets an array, which requires further processing.
Other usage
echo $crawler->filterXPath(‘//body/p‘)->attr(‘class‘);
You can get the value "message" of the class attribute corresponding to the first p tag
$crawler->filterXPath(‘//p[@class="样式"]‘)->filter(‘a‘)->attr(‘href‘); $crawler->filterXPath(‘//p[@class="样式"]‘)->filter(‘a>img‘)->extract(array(‘alt‘, ‘href‘))
and above They are some methods of obtaining tag attributes.
filter is different from filter to try.
Generally speaking, I feel that DomCrawler is easier to use than simple html dom. Maybe it is because I use it more easily.
The above are just the basic functions of Crawler. For more usage, please refer to the functions in the Crawler part of the symfony manual
http://api.symfony.com/3.2/Symfony/Component/DomCrawler/Crawler .html
The main problem with Crawler is that there are too few examples. There are no usage examples in the function manual, so you can only explore it in actual use. . . .
symfony's documentation about DomCrawler, there are a few examples
http://symfony.com/doc/current/components/dom_crawler.html
The above is the detailed content of Using Symfony's Crawler component to analyze HTML_php instances in laravel. For more information, please follow other related articles on the PHP Chinese website!

Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

Video Face Swap
Swap faces in any video effortlessly with our completely free AI face swap tool!

Hot Article

Hot Tools

Notepad++7.3.1
Easy-to-use and free code editor

SublimeText3 Chinese version
Chinese version, very easy to use

Zend Studio 13.0.1
Powerful PHP integrated development environment

Dreamweaver CS6
Visual web development tools

SublimeText3 Mac version
God-level code editing software (SublimeText3)

Hot Topics



Laravel - Artisan Commands - Laravel 5.7 comes with new way of treating and testing new commands. It includes a new feature of testing artisan commands and the demonstration is mentioned below ?

Laravel - Pagination Customizations - Laravel includes a feature of pagination which helps a user or a developer to include a pagination feature. Laravel paginator is integrated with the query builder and Eloquent ORM. The paginate method automatical

Method for obtaining the return code when Laravel email sending fails. When using Laravel to develop applications, you often encounter situations where you need to send verification codes. And in reality...

Laravel schedule task run unresponsive troubleshooting When using Laravel's schedule task scheduling, many developers will encounter this problem: schedule:run...

The method of handling Laravel's email failure to send verification code is to use Laravel...

How to implement the table function of custom click to add data in dcatadmin (laravel-admin) When using dcat...

The impact of sharing of Redis connections in Laravel framework and select methods When using Laravel framework and Redis, developers may encounter a problem: through configuration...

Laravel - Dump Server - Laravel dump server comes with the version of Laravel 5.7. The previous versions do not include any dump server. Dump server will be a development dependency in laravel/laravel composer file.
