This article mainly introduces the related information of MySQL deduplication method. Friends who need it can refer to
MySQL deduplication method
[Elementary] There are very few duplicate lines
Use distinct to find them, and then manually delete them one by one.
[Intermediate] Deduplication according to the repetition of a single field
For example: Deduplication of the id field
Usage: Get the id For the values of duplicate fields, use the rows where the same id field is located to compare the fields with different data, and delete all duplicate rows except the row where the smallest (or largest) field is located. Generally, the primary key is used for comparison, because the value of the primary key must be a unique value and must not be the same.
id name 1 a 1 b 2 c 2 a 3 c
Result:
id name 1 a 2 a
Operation:
delete from a_tmp where id in (select * from (select b.id from a_tmp b group by b.id having count(b.id) >1) bb) and name not in (select * from (select min(a.name) from a_tmp a GROUP BY a.id having count(a.id) >1) aa);
Note:
The above bold and green words must be aliased and must use the format select * from (...), otherwise an error will be reported:
[Err] 1093 - You can't specify target table 'a_tmp ' for update in FROM clause
[Advanced] Repeat by multiple fields
For example: the same id and name Deduplication, that is: rows with the same ID and name are counted as duplicate rows, rows with the same ID but different names are counted as non-duplicate rows
Usage method: similar to a single field, generally use the primary key To compare, because the value of the primary key must be a unique value.
id name rowid 1 a 1 1 a 2 1 b 3 2 b 4 2 b 5 3 c 6 3 d 7
Result:
id name rowid 1 a 1 1 b 3 2 b 4 3 c 6 3 d 7
Operation:
First type:
delete from a_tmp where (id,name) in (select * from (select b.id,b.name from a_tmp b group by b.id,b.name having count(b.id) >1) bb) and rowid not in (select * from (select min(a.rowid) from a_tmp a group by a.id,a.name having count(a.id) >1) aa);
Second type :
Connect the values of the id and name fields and insert them into the temporary table b_tmp, so that you can use the [Intermediate] single field judgment deletion method.
#Insert the value of the connection between the two fields and the unique value field in the a_tmp table into the b_tmp table
insert into b_tmp select concat(id,name),rowid from a_tmp; #查出需要留下来的行 select id_name,max(rowid) from b_tmp group by id_name having count(id_name)>1; #使用【中级】的方法,或存储过程完成去重的工作
[Ultimate] Each row has two copies of the same data
For example:
Instructions for use: The entire row of data is the same and cannot be deleted using SQL statements because there is no conditional restriction that can be used to leave one row and delete all rows that are identical to it. . There are no different fields. You can create different fields by yourself, that is: add a field, set it to auto-increment, and set it as the primary key, and it will automatically add the upper value.
id name 1 a 1 a 1 b 1 b 2 c 2 c 3 c 3 c
Result:
id name rowid 1 a 1 1 b 3 2 c 5 3 c 7
Operation:
Add a self-increasing field and temporarily set it as the primary key.
Use the [Intermediate] and [Advanced] methods above.
The above is the detailed content of Mysql deduplication method. For more information, please follow other related articles on the PHP Chinese website!

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了关于架构原理的相关内容,MySQL Server架构自顶向下大致可以分网络连接层、服务层、存储引擎层和系统文件层,下面一起来看一下,希望对大家有帮助。

mysql的msi与zip版本的区别:1、zip包含的安装程序是一种主动安装,而msi包含的是被installer所用的安装文件以提交请求的方式安装;2、zip是一种数据压缩和文档存储的文件格式,msi是微软格式的安装包。

方法:1、利用right函数,语法为“update 表名 set 指定字段 = right(指定字段, length(指定字段)-1)...”;2、利用substring函数,语法为“select substring(指定字段,2)..”。

在mysql中,可以利用char()和REPLACE()函数来替换换行符;REPLACE()函数可以用新字符串替换列中的换行符,而换行符可使用“char(13)”来表示,语法为“replace(字段名,char(13),'新字符串') ”。

转换方法:1、利用cast函数,语法“select * from 表名 order by cast(字段名 as SIGNED)”;2、利用“select * from 表名 order by CONVERT(字段名,SIGNED)”语句。

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了关于MySQL复制技术的相关问题,包括了异步复制、半同步复制等等内容,下面一起来看一下,希望对大家有帮助。

在mysql中,可以利用REGEXP运算符判断数据是否是数字类型,语法为“String REGEXP '[^0-9.]'”;该运算符是正则表达式的缩写,若数据字符中含有数字时,返回的结果是true,反之返回的结果是false。

在mysql中,是否需要commit取决于存储引擎:1、若是不支持事务的存储引擎,如myisam,则不需要使用commit;2、若是支持事务的存储引擎,如innodb,则需要知道事务是否自动提交,因此需要使用commit。


Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

AI Hentai Generator
Generate AI Hentai for free.

Hot Article

Hot Tools

VSCode Windows 64-bit Download
A free and powerful IDE editor launched by Microsoft

SublimeText3 Mac version
God-level code editing software (SublimeText3)

EditPlus Chinese cracked version
Small size, syntax highlighting, does not support code prompt function

MantisBT
Mantis is an easy-to-deploy web-based defect tracking tool designed to aid in product defect tracking. It requires PHP, MySQL and a web server. Check out our demo and hosting services.

mPDF
mPDF is a PHP library that can generate PDF files from UTF-8 encoded HTML. The original author, Ian Back, wrote mPDF to output PDF files "on the fly" from his website and handle different languages. It is slower than original scripts like HTML2FPDF and produces larger files when using Unicode fonts, but supports CSS styles etc. and has a lot of enhancements. Supports almost all languages, including RTL (Arabic and Hebrew) and CJK (Chinese, Japanese and Korean). Supports nested block-level elements (such as P, DIV),
