search
HomeDatabaseMysql TutorialHow to implement random extraction in MySQL

1. Introduction

Now there is a requirement to randomly select three words at a time from a word list.

The table creation statement of this table is as follows:

mysql> Create table 'words'(
    'id' int(11) not null auto_increment;
    'word' varchar(64) default null;
    primary key ('id')
) ENGINE=InnoDB;

Then we insert 10,000 rows of data into it. Next let's see how to randomly select 3 words from it.

2. Memory temporary table

First of all, we usually think of using order by rand() to implement this logic:

mysql> select word from words order by rand() limit 3;

Although this sentence is very simple, but the execution The process is more complicated. We use explain to see the execution of the statement:

How to implement random extraction in MySQL

Using temporary in the Extra field indicates that a temporary table needs to be used, and Using filesort indicates that sorting is required. That is to say, sorting operation is required.

For InnoDB tables, performing full field sorting can reduce disk access, so it will be preferred.

How to implement random extraction in MySQL

#For memory tables, the table return process simply directly accesses the memory to obtain the data based on the location of the data rows, and does not result in multiple disk accesses at all. . So at this time MySQL will give priority to rowid sorting.

How to implement random extraction in MySQL

Let’s sort out the execution process of this statement:

  • Create a temporary table, this table Using the memory engine, there are two fields in the table. The first field is of double type, marked as R, and the second field is of varchar(64) type, marked as W. And this table has no index.

  • From the words table, take out all words in primary key order. For each word, call the rand() function to randomly generate a random decimal number greater than 0 and less than 1, and store the random decimal number and the word in the R and W fields of the temporary table respectively.

  • The next step is to sort according to field R

  • Initialize sort_buffer. sort_buffer includes a double type and an integer field.

  • Take out the R value and position information row by row from the temporary memory table, and store them in the two fields of sort_buffer respectively.

  • sort_buffer is sorted according to the R value

  • After the sorting is completed, the location information of the first three results is taken out, and the corresponding information is taken out from the memory temporary table The word is returned to the client.

The process diagram is as follows:

How to implement random extraction in MySQL

The location information mentioned above is actually the location of the row, that is, This is the rowid we mentioned before.

For the InnoDB engine, there are two processing methods for tables with or without primary keys:

  • For InnoDB tables with primary keysFor example, this rowid is the primary key id

  • For InnoDB tables without primary keys, this rowid is generated by the system and is used to identify different rows .

Therefore, order by randn() uses a memory temporary table, and the sorting method of the memory temporary table uses the rowid sorting method.

3. Disk temporary table

Not all temporary tables are memory temporary tables. The tmp_table_size configuration limits the size of the memory temporary table. If this size is exceeded, the disk temporary table will be used. The InnoDB engine uses disk temporary tables by default.

4. Priority queue sorting algorithm

After MySQL5.6, the priority queue sorting algorithm was introduced. This algorithm does not require the use of temporary files. The original merge sort algorithm requires the use of temporary files.

Because when you use the merge algorithm, you actually only need to get the top 3, but if you run out of merge sort, the whole thing is already in order, causing a waste of resources.

The priority queue sorting algorithm can only take the top three. The execution process is as follows:

  • For these 10,000 (R, rowid) to be sorted, the top three are taken first. Three rows are constructed into a heap, and the largest value is placed on the top of the heap;

  • Take the next row (R’, rowid’) and compare it with the largest R in the current heap. If R’ is less than R, remove (R, rowid) from the heap and replace it with (R’, rowid’).

  • Repeat the above process.

The process is shown in the figure below:

How to implement random extraction in MySQL

But when the limit number is relatively large, it is more difficult to maintain the heap, so it will Use the merge sort algorithm.

The above is the detailed content of How to implement random extraction in MySQL. For more information, please follow other related articles on the PHP Chinese website!

Statement
This article is reproduced at:亿速云. If there is any infringement, please contact admin@php.cn delete
图文详解mysql架构原理图文详解mysql架构原理May 17, 2022 pm 05:54 PM

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了关于架构原理的相关内容,MySQL Server架构自顶向下大致可以分网络连接层、服务层、存储引擎层和系统文件层,下面一起来看一下,希望对大家有帮助。

mysql怎么替换换行符mysql怎么替换换行符Apr 18, 2022 pm 03:14 PM

在mysql中,可以利用char()和REPLACE()函数来替换换行符;REPLACE()函数可以用新字符串替换列中的换行符,而换行符可使用“char(13)”来表示,语法为“replace(字段名,char(13),'新字符串') ”。

mysql的msi与zip版本有什么区别mysql的msi与zip版本有什么区别May 16, 2022 pm 04:33 PM

mysql的msi与zip版本的区别:1、zip包含的安装程序是一种主动安装,而msi包含的是被installer所用的安装文件以提交请求的方式安装;2、zip是一种数据压缩和文档存储的文件格式,msi是微软格式的安装包。

mysql怎么去掉第一个字符mysql怎么去掉第一个字符May 19, 2022 am 10:21 AM

方法:1、利用right函数,语法为“update 表名 set 指定字段 = right(指定字段, length(指定字段)-1)...”;2、利用substring函数,语法为“select substring(指定字段,2)..”。

mysql怎么将varchar转换为int类型mysql怎么将varchar转换为int类型May 12, 2022 pm 04:51 PM

转换方法:1、利用cast函数,语法“select * from 表名 order by cast(字段名 as SIGNED)”;2、利用“select * from 表名 order by CONVERT(字段名,SIGNED)”语句。

MySQL复制技术之异步复制和半同步复制MySQL复制技术之异步复制和半同步复制Apr 25, 2022 pm 07:21 PM

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了关于MySQL复制技术的相关问题,包括了异步复制、半同步复制等等内容,下面一起来看一下,希望对大家有帮助。

带你把MySQL索引吃透了带你把MySQL索引吃透了Apr 22, 2022 am 11:48 AM

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了mysql高级篇的一些问题,包括了索引是什么、索引底层实现等等问题,下面一起来看一下,希望对大家有帮助。

mysql怎么判断是否是数字类型mysql怎么判断是否是数字类型May 16, 2022 am 10:09 AM

在mysql中,可以利用REGEXP运算符判断数据是否是数字类型,语法为“String REGEXP '[^0-9.]'”;该运算符是正则表达式的缩写,若数据字符中含有数字时,返回的结果是true,反之返回的结果是false。

See all articles

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

AI Hentai Generator

AI Hentai Generator

Generate AI Hentai for free.

Hot Article

Repo: How To Revive Teammates
1 months agoBy尊渡假赌尊渡假赌尊渡假赌
R.E.P.O. Energy Crystals Explained and What They Do (Yellow Crystal)
2 weeks agoBy尊渡假赌尊渡假赌尊渡假赌
Hello Kitty Island Adventure: How To Get Giant Seeds
1 months agoBy尊渡假赌尊渡假赌尊渡假赌

Hot Tools

Dreamweaver Mac version

Dreamweaver Mac version

Visual web development tools

Atom editor mac version download

Atom editor mac version download

The most popular open source editor

WebStorm Mac version

WebStorm Mac version

Useful JavaScript development tools

VSCode Windows 64-bit Download

VSCode Windows 64-bit Download

A free and powerful IDE editor launched by Microsoft

Notepad++7.3.1

Notepad++7.3.1

Easy-to-use and free code editor