


Ideas and sample codes for using PHP instead of JS to play with DOM_PHP Tutorial
The origin of the matter is relatively simple. I need to organize the data of a navigation page and write it into the database. A more intuitive method is to analyze the HTML file. The common method is to use PHP regular expressions to match. However, it is difficult to develop and maintain in this way, and the code readability is very poor.
The data on the navigation page is regularly arranged in the DOM tree. It can be easily operated with several loops using JS. Moreover, JS needs to rely on the browser and it is difficult to operate the database. In fact, PHP has a ready-made class library to add, delete, modify and check nodes in the DOM tree. I will make some notes here.
There are two classes involved here: DOMDocument and DOMXPath.
In fact, the idea is relatively clear, which is to convert an html file into the data structure of a DOM tree through DOMDocument, and then use an instance of DOMXPath to search the DOM tree to get the specific node you want, and then you can Traverse the subtree of the node to get the desired result.
There is such a navigation html file "./hao.html" in the current directory
Now we need to get the Chinese content of all tags, the php code is as follows:
//Convert html/xml file into DOM tree
$dom = new DOMDocument();
$dom->loadHTMLFile("hao.html");
//Get all dl tags with class fix
// example 1: for everything with an id
//$elements = $xpath->query("//*[@id]");
// example 2: for node data in a selected id
//$elements = $xpath->query("/html/body/div[@id='yourTagIdHere']");
// example 3: same as above with wildcard
//$elements = $xpath->query("*/div[@id='yourTagIdHere']");
$xpath = new DOMXPath($dom);
$dls = $xpath->query('//dl[@class="fix"]');
foreach ($dls as $dl) {
$spans = $dl->childNodes;
foreach ($spans as $span) {
echo trim($span->textContent)."t";
}
echo "n";
}
? >
The output result is as follows:
Note: It is worth noting that the default encoding method of DOMDocument is Latin, so when processing UTF-encoded Chinese, you need to enter < head> followed by
in other locations, or just write both It's not recognized

Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

AI Hentai Generator
Generate AI Hentai for free.

Hot Article

Hot Tools

Notepad++7.3.1
Easy-to-use and free code editor

SublimeText3 Chinese version
Chinese version, very easy to use

Zend Studio 13.0.1
Powerful PHP integrated development environment

Dreamweaver CS6
Visual web development tools

SublimeText3 Mac version
God-level code editing software (SublimeText3)

Hot Topics



In this chapter, we will understand the Environment Variables, General Configuration, Database Configuration and Email Configuration in CakePHP.

PHP 8.4 brings several new features, security improvements, and performance improvements with healthy amounts of feature deprecations and removals. This guide explains how to install PHP 8.4 or upgrade to PHP 8.4 on Ubuntu, Debian, or their derivati

To work with date and time in cakephp4, we are going to make use of the available FrozenTime class.

Working with database in CakePHP is very easy. We will understand the CRUD (Create, Read, Update, Delete) operations in this chapter.

To work on file upload we are going to use the form helper. Here, is an example for file upload.

In this chapter, we are going to learn the following topics related to routing ?

CakePHP is an open-source framework for PHP. It is intended to make developing, deploying and maintaining applications much easier. CakePHP is based on a MVC-like architecture that is both powerful and easy to grasp. Models, Views, and Controllers gu

Validator can be created by adding the following two lines in the controller.
