How to solve the problem of count distinct multiple columns in mysql-Mysql Tutorial-php.cn

Home

Database

Mysql Tutorial

How to solve the problem of count distinct multiple columns in mysql

王林

Jun 03, 2023 am 10:49 AM

mysqlcountdistinct

The reproduced test database is as follows:

CREATE TABLE `test_distinct` (
  `id` int(11) NOT NULL AUTO_INCREMENT,
  `a` varchar(50) CHARACTER SET utf8 DEFAULT NULL,
  `b` varchar(50) CHARACTER SET utf8 DEFAULT NULL,
  PRIMARY KEY (`id`)
) ENGINE=InnoDB AUTO_INCREMENT=1 DEFAULT CHARSET=latin1;

The test data in the table is as follows. Now we need to count the number of columns after deduplication of these three columns.

How to solve the problem of count distinct multiple columns in mysql

Problem Analysis

My friend gave me four query statements to locate the problem

SELECT COUNT(*) AS cnt FROM test_distinct;
SELECT COUNT(DISTINCT id, a, b) as cnt FROM test_distinct;
SELECT id, a, b, COUNT(*) AS cnt FROM test_distinct GROUP BY id, a, b HAVING cnt > 1;
SELECT 
	l.id AS l_id,
	l.a AS l_a,
	l.b AS l_b,
	r.id AS r_id,
	r.a AS r_a,
	r.b AS r_b
FROM test_distinct l LEFT JOIN test_distinct r
ON l.id = r.id AND l.a = r.a AND l.b = r.b
WHERE r.id is NULL or r.id = &#39;null&#39;;

The query results are as follows:

How to solve the problem of count distinct multiple columns in mysql

Notice! ! ! From the test data, we can quickly guess where the problem lies, but it turns out that there are more than 30,000 pieces of data in the table, and it is impossible to view the data with the naked eye.

There are two counterintuitive points in the above query results:

The second piece of data is missing after deduplication statistics, but the result of the third piece of data shows There is no identical data.
When using the same table to do a left outer connection, the driving table has data, but the driven table is empty.

Let’s look at the second question first. The official document has the following explanation:

When using the ON clause, the conditions it contains The expression is the same as that used in the WHERE clause. A common situation is to use the ON clause to specify the join conditions of the table, and use the WHERE clause to limit the rows included in the result set.
If there are no matching rows in the right table for the conditions in the ON or USING part of the LEFT JOIN, then the right table uses all columns set to NULL.
You cannot use arithmetic comparison operators (such as =, ) to compare NULL.

SELECT NULL = NULL;
SELECT NULL IS NULL;

How to solve the problem of count distinct multiple columns in mysql

So the second problem is that the result of NULL=NULL is always False, which results in the two rows originally Equal data results are not equal.

But this does not solve the first problem: why a piece of data disappeared after deduplication. However, we can guess that the missing data is probably related to the NULL value.

We separate the two operations of count and distinct:

SELECT COUNT(*) as cnt FROM (SELECT  DISTINCT id, a, b FROM test_distinct) as tmp;

How to solve the problem of count distinct multiple columns in mysql

Huh? The result is correct, which means that the query plan generated by count(distinct expr) may be different from what we imagined. It is not to remove duplicates first and then count. Use explain to analyze the query plan of the two statements. As shown below:

How to solve the problem of count distinct multiple columns in mysql

As you can see from the table, the mysql execution engine directly counts count(distinct expr)As a query, check the official documentation:

How to solve the problem of count distinct multiple columns in mysql

Solution

The problem has finally been clarified. There are two ways to solve this problem. The first is to remove duplicates first and then count. The second is to use the IFNULL() function:

SELECT COUNT(DISTINCT id, a, IFNULL(b, &#39;0&#39;)) as cnt FROM test_distinct;

In addition, count( )Use:

SELECT id, a, b, COUNT(*) FROM test_distinct GROUP BY id, a, b;
SELECT id, a, b, COUNT(b) FROM test_distinct GROUP BY id, a, b;

How to solve the problem of count distinct multiple columns in mysql

Knowledge point

You cannot use arithmetic comparison operators (such as =, ) to compare null values;
count(distinct expr) returns the number of distinct and non-empty rows in the expr column;
COUNT() has two distinct uses: it can be used to count the number of values in a column, or it can be used to count the number of rows. When counting column values, the column value is required to be non-empty (NULL is not counted). When a column or expression is specified in parentheses of the COUNT() function, the function counts the number of results that have a value in the expression. Another function of COUNT() is to count the number of rows in the result set. When MySQL confirms that the expression value within the parentheses cannot be empty, it is actually counting the number of rows. The simplest thing is when we use COUNT(). In this case, the wildcard does not expand to all columns as we guessed. In fact, it will ignore all columns and directly count all rows - "High-Performance MySQL";
In InnoDB, SELECT COUNT(*) and SELECT COUNT(1) are processed in the same way, and there is no performance difference.

The above is the detailed content of How to solve the problem of count distinct multiple columns in mysql. For more information, please follow other related articles on the PHP Chinese website!

Statement

This article is reproduced at:亿速云. If there is any infringement, please contact admin@php.cn delete

How Do I Drop or Modify an Existing View in MySQL?May 16, 2025 am 12:11 AM

TodropaviewinMySQL,use"DROPVIEWIFEXISTSview_name;"andtomodifyaview,use"CREATEORREPLACEVIEWview_nameASSELECT...".Whendroppingaview,considerdependenciesanduse"SHOWCREATEVIEWview_name;"tounderstanditsstructure.Whenmodifying

MySQL Views: Which design patterns can I use with it?May 16, 2025 am 12:10 AM

MySQLViewscaneffectivelyutilizedesignpatternslikeAdapter,Decorator,Factory,andObserver.1)AdapterPatternadaptsdatafromdifferenttablesintoaunifiedview.2)DecoratorPatternenhancesdatawithcalculatedfields.3)FactoryPatterncreatesviewsthatproducedifferentda

What Are the Advantages of Using Views in MySQL?May 16, 2025 am 12:09 AM

ViewsinMySQLarebeneficialforsimplifyingcomplexqueries,enhancingsecurity,ensuringdataconsistency,andoptimizingperformance.1)Theysimplifycomplexqueriesbyencapsulatingthemintoreusableviews.2)Viewsenhancesecuritybycontrollingdataaccess.3)Theyensuredataco

How Can I Create a Simple View in MySQL?May 16, 2025 am 12:08 AM

TocreateasimpleviewinMySQL,usetheCREATEVIEWstatement.1)DefinetheviewwithCREATEVIEWview_nameAS.2)SpecifytheSELECTstatementtoretrievedesireddata.3)Usetheviewlikeatableforqueries.Viewssimplifydataaccessandenhancesecurity,butconsiderperformance,updatabil

MySQL Create User Statement: Examples and Common ErrorsMay 16, 2025 am 12:04 AM

TocreateusersinMySQL,usetheCREATEUSERstatement.1)Foralocaluser:CREATEUSER'localuser'@'localhost'IDENTIFIEDBY'securepassword';2)Foraremoteuser:CREATEUSER'remoteuser'@'%'IDENTIFIEDBY'strongpassword';3)Forauserwithaspecifichost:CREATEUSER'specificuser'@

What Are the Limitations of Using Views in MySQL?May 14, 2025 am 12:10 AM

MySQLviewshavelimitations:1)Theydon'tsupportallSQLoperations,restrictingdatamanipulationthroughviewswithjoinsorsubqueries.2)Theycanimpactperformance,especiallywithcomplexqueriesorlargedatasets.3)Viewsdon'tstoredata,potentiallyleadingtooutdatedinforma

Securing Your MySQL Database: Adding Users and Granting PrivilegesMay 14, 2025 am 12:09 AM

ProperusermanagementinMySQLiscrucialforenhancingsecurityandensuringefficientdatabaseoperation.1)UseCREATEUSERtoaddusers,specifyingconnectionsourcewith@'localhost'or@'%'.2)GrantspecificprivilegeswithGRANT,usingleastprivilegeprincipletominimizerisks.3)

What Factors Influence the Number of Triggers I Can Use in MySQL?May 14, 2025 am 12:08 AM

MySQLdoesn'timposeahardlimitontriggers,butpracticalfactorsdeterminetheireffectiveuse:1)Serverconfigurationimpactstriggermanagement;2)Complextriggersincreasesystemload;3)Largertablesslowtriggerperformance;4)Highconcurrencycancausetriggercontention;5)M

See all articles

Hot AI Tools

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress images for free

Clothoff.io

AI clothes remover

Video Face Swap

Swap faces in any video effortlessly with our completely free AI face swap tool!

Hot Article

Roblox: Grow A Garden - Complete Mutation Guide

4 weeks agoByDDD

Roblox: Bubble Gum Simulator Infinity - How To Get And Use Royal Keys

4 weeks agoBy尊渡假赌尊渡假赌尊渡假赌

Nordhold: Fusion System, Explained

1 months agoBy尊渡假赌尊渡假赌尊渡假赌

Mandragora: Whispers Of The Witch Tree - How To Unlock The Grappling Hook

4 weeks agoBy尊渡假赌尊渡假赌尊渡假赌

Clair Obscur: Expedition 33 - How To Get Perfect Chroma Catalysts

2 weeks agoBy尊渡假赌尊渡假赌尊渡假赌

Hot Tools

Safe Exam Browser

Safe Exam Browser is a secure browser environment for taking online exams securely. This software turns any computer into a secure workstation. It controls access to any utility and prevents students from using unauthorized resources.

SublimeText3 English version

Recommended: Win version, supports code prompts!

MinGW - Minimalist GNU for Windows

This project is in the process of being migrated to osdn.net/projects/mingw, you can continue to follow us there. MinGW: A native Windows port of the GNU Compiler Collection (GCC), freely distributable import libraries and header files for building native Windows applications; includes extensions to the MSVC runtime to support C99 functionality. All MinGW software can run on 64-bit Windows platforms.

mPDF

mPDF is a PHP library that can generate PDF files from UTF-8 encoded HTML. The original author, Ian Back, wrote mPDF to output PDF files "on the fly" from his website and handle different languages. It is slower than original scripts like HTML2FPDF and produces larger files when using Unicode fonts, but supports CSS styles etc. and has a lot of enhancements. Supports almost all languages, including RTL (Arabic and Hebrew) and CJK (Chinese, Japanese and Korean). Supports nested block-level elements (such as P, DIV),

Dreamweaver CS6

Visual web development tools

Hot Topics

1677

1430

1333

1278

1257