Home Backend Development PHP Tutorial Detailed example explanation of PHP+Sphinx+Mysql development of search engine

Detailed example explanation of PHP+Sphinx+Mysql development of search engine

Feb 10, 2018 am 10:39 AM
php search engine

When everyone hears about search engines, they will find it difficult to write and have no idea at all. In fact, PHP can also be used for search engine development, but PHP needs to be combined with sphinx and mysql to develop the search engine we want. I want to know about PHP How to carry out search engine development! Let’s take a look! !

First we download the Sphinx tool, download address official website download address: www.sphinxsearch.com, find sphinx-2.2.10-release-win64.zip, download this for 64-bit, unzip it to us Under the PHP running directory, it is convenient to view the results on subsequent web pages.

sphinx introduction:

Sphinx is the abbreviation of SQL Phrase Index (query phrase index). Sphinx is a full-text search engine based on SQL. The API interfaces it provides include :PHP, Python, Perl, Ruby, java, etc. At the same time, an engine plug-in SphinxSE is designed for MySQL, which is a distributed full-text retrieval system.
Advantages:
High-speed indexing can reach 10M/s
High-performance search (text data in 2-4G On average, the average response time for each retrieval is less than 0.1 seconds)
Can handle massive amounts of data (currently known to be able to process 100G of text data, and 100M of documents on a single CPU system)
Provides excellent correlation algorithm, composite ranking method based on phrase similarity and statistical BM2
Supports distributed search
Provide document fragment generation function
Can be used as a Mysql storage engine to provide search services
Support Boolean, phrase, word similarity and other search modes
Disadvantages:
Must have a primary key
The primary key must be an integer
Not responsible for data storage
The configuration is not flexible

The sphinx structure after decompression is as shown in the figure:


The following is our process For related configuration, see sphinx-min.conf.in in the picture, copy it to our bin directory for easy use and change the name to sphinx.conf,

Modify the content inside:

source src1
{
	type			= mysql

	sql_host		= localhost #主机地址
	sql_user		= root#帐号
	sql_pass		=     #密码
	sql_db			= sphinx  #数据库
	sql_port		= 3306	# 数据库端口 3306
	sql_query		= SELECT id, name, age FROM users #查询语句
	sql_attr_uint		= group_id
	sql_attr_timestamp	= date_added
	sql_query_pre = set names utf8   #数据库编码
}


index test1
{
	source			= src1
	path			= D:/myapaphe/www/sphinx/data #这个一定要配置
	charset_type = utf-8 #指定编码
	ngram_len = 1        #要找中文需指定为1.
	ngram_chars = U+3000..U+2FA1F
	
}

indexer
{
	mem_limit		= 128M
}
searchd
{
	listen			= 9312
	listen			= 9306:mysql41
	log			= D:\myapaphe\www\sphinx\log\searchd.log  #进程日志
	query_log		= D:\myapaphe\www\sphinx\log\query.log    #查询日志

	read_timeout		= 5
	max_children		= 30
	pid_file		= D:\myapaphe\www\sphinx\log\searchd.pid 
	seamless_rotate		= 1
	preopen_indexes		= 1
	unlink_old		= 1
	workers			= threads # for RT to work
	binlog_path		= D:\myapaphe\www\sphinx\data
}
Copy after login

The above must be configured, and the path must match your own path.

Next generate the query index:


Install searchd service:


Next load the configuration file:


##Start the service:


OK the previous configuration work and service startup have been completed. Now start the code:

Create test3.php under the api folder under sphinx and run test3.php

<?php 
require ( "sphinxapi.php" );
$s = new SphinxClient();
$s->SetServer(&#39;localhost&#39;,9312);
$result = $s->Query(&#39;高七&#39;);
echo &#39;<pre class="brush:php;toolbar:false">&#39;;
print_r($result);
Copy after login


Garbled characters are because cmd defaults to gbk encoding. Let’s put it in the browser to view:


We see that sphinx does not find the complete result but returns the ID to us, allowing us to check the data based on the ID.

The following is a query time comparison:


The time I tested on more than 40,000 pieces of data was 0.001s. Let’s take a look at mysql How long does the query take:


We see that it takes 0.04s, there is not much data, and the result is not that obvious, but the gap of 0.039s is not small.

This completes the integration of sphinx, I hope it can help everyone.

related suggestion:

php Detailed explanation of calling existing search engines

php Function code to determine whether the visitor is a search engine spider

The above is the detailed content of Detailed example explanation of PHP+Sphinx+Mysql development of search engine. For more information, please follow other related articles on the PHP Chinese website!

Statement of this Website
The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

AI Hentai Generator

AI Hentai Generator

Generate AI Hentai for free.

Hot Article

R.E.P.O. Energy Crystals Explained and What They Do (Yellow Crystal)
3 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
R.E.P.O. Best Graphic Settings
3 weeks ago By 尊渡假赌尊渡假赌尊渡假赌
R.E.P.O. How to Fix Audio if You Can't Hear Anyone
3 weeks ago By 尊渡假赌尊渡假赌尊渡假赌

Hot Tools

Notepad++7.3.1

Notepad++7.3.1

Easy-to-use and free code editor

SublimeText3 Chinese version

SublimeText3 Chinese version

Chinese version, very easy to use

Zend Studio 13.0.1

Zend Studio 13.0.1

Powerful PHP integrated development environment

Dreamweaver CS6

Dreamweaver CS6

Visual web development tools

SublimeText3 Mac version

SublimeText3 Mac version

God-level code editing software (SublimeText3)

CakePHP Project Configuration CakePHP Project Configuration Sep 10, 2024 pm 05:25 PM

In this chapter, we will understand the Environment Variables, General Configuration, Database Configuration and Email Configuration in CakePHP.

PHP 8.4 Installation and Upgrade guide for Ubuntu and Debian PHP 8.4 Installation and Upgrade guide for Ubuntu and Debian Dec 24, 2024 pm 04:42 PM

PHP 8.4 brings several new features, security improvements, and performance improvements with healthy amounts of feature deprecations and removals. This guide explains how to install PHP 8.4 or upgrade to PHP 8.4 on Ubuntu, Debian, or their derivati

CakePHP Date and Time CakePHP Date and Time Sep 10, 2024 pm 05:27 PM

To work with date and time in cakephp4, we are going to make use of the available FrozenTime class.

CakePHP Working with Database CakePHP Working with Database Sep 10, 2024 pm 05:25 PM

Working with database in CakePHP is very easy. We will understand the CRUD (Create, Read, Update, Delete) operations in this chapter.

CakePHP File upload CakePHP File upload Sep 10, 2024 pm 05:27 PM

To work on file upload we are going to use the form helper. Here, is an example for file upload.

CakePHP Routing CakePHP Routing Sep 10, 2024 pm 05:25 PM

In this chapter, we are going to learn the following topics related to routing ?

Discuss CakePHP Discuss CakePHP Sep 10, 2024 pm 05:28 PM

CakePHP is an open-source framework for PHP. It is intended to make developing, deploying and maintaining applications much easier. CakePHP is based on a MVC-like architecture that is both powerful and easy to grasp. Models, Views, and Controllers gu

CakePHP Creating Validators CakePHP Creating Validators Sep 10, 2024 pm 05:26 PM

Validator can be created by adding the following two lines in the controller.

See all articles