MySQL批量插入遇上唯一索引避免方法_MySQL

MySQL批量插入遇上唯一索引避免方法_MySQL

WBOYWBOYWBOYWBOYWBOYWBOYWBOYWBOYWBOYWBOYWBOYWBOYWB

Jun 01, 2016 pm 01:23 PM

bitsCN.com

一、背景

以前使用SQL Server进行表分区的时候就碰到很多关于唯一索引的问题：Step8：SQL Server 当表分区遇上唯一约束，没想到在MySQL的分区中一样会遇到这样的问题：MySQL表分区实战。

今天我们来了解MySQL唯一索引的一些知识：包括如何创建，如何批量插入，还有一些技巧上SQL；

这些问题的根源在什么地方？有什么共同点？MySQL中也有分区对齐的概念？唯一索引是在很多系统中都会出现的要求，有什么办法可以避免？它对性能的影响有多大？

二、过程

(一) 导入差异数据，忽略重复数据，IGNORE INTO的使用

在MySQL创建表的时候，我们通常创建一个表的时候是以一个自增ID值作为主键，那么MySQL就会以PRIMARY KEY作为聚集索引键和主键，既然是主键，那当然是唯一的了，所以重复执行下面的插入语句会报1062错误：如Figure1所示；

-- 创建测试表
CREATE TABLE `testtable` (
`Id` INT(11) UNSIGNED NOT NULL AUTO_INCREMENT,
`UserId` INT(11) DEFAULT NULL,
`UserName` VARCHAR(10) DEFAULT NULL,
`UserType` INT(11) DEFAULT NULL,
PRIMARY KEY (`Id`)
) ENGINE=INNODB DEFAULT CHARSET=utf8;

-- 插入测试数据
INSERT INTO testtable(Id,UserId,UserName,UserType)
VALUES(1,101,'aa',1),(2,102,'bbb',2),(3,103,'ccc',3);

u1_1062

（Figure1：Duplicate entry '1' for key 'PRIMARY'）

但是在实际的生产环境中，需求往往是需要在UserId键值中设置唯一索引，今天我就以这个作为示例，进行唯一索引的测试：

-- 创建测试表1
CREATE TABLE `testtable1` (
`Id` INT(11) UNSIGNED NOT NULL AUTO_INCREMENT,
`UserId` INT(11) DEFAULT NULL,
`UserName` VARCHAR(10) DEFAULT NULL,
`UserType` INT(11) DEFAULT NULL,
PRIMARY KEY (`Id`),
UNIQUE KEY `IX_UserId` (`UserId`)
) ENGINE=INNODB DEFAULT CHARSET=utf8;

-- 创建测试表2
CREATE TABLE `testtable2` (
`Id` INT(11) UNSIGNED NOT NULL AUTO_INCREMENT,
`UserId` INT(11) DEFAULT NULL,
`UserName` VARCHAR(10) DEFAULT NULL,
`UserType` INT(11) DEFAULT NULL,
PRIMARY KEY (`Id`),
UNIQUE KEY `IX_UserId` (`UserId`)
) ENGINE=INNODB DEFAULT CHARSET=utf8;

-- 插入测试数据1
INSERT INTO testtable1(Id,UserId,UserName,UserType)
VALUES(1,101,'aa',1),(2,102,'bbb',2),(3,103,'ccc',3);

-- 插入测试数据2
INSERT INTO testtable2(Id,UserId,UserName,UserType)
VALUES(1,201,'aaa',1),(2,202,'bbb',2),(3,203,'ccc',3),(4,101,'xxxx',5);

u2_table1

（Figure2：testtable1记录）

u3_table2

（Figure3：testtable2记录）

通过执行上面的SQL脚本，我们在testtable1和testtable2都创建了唯一索引：UNIQUE KEY `IX_UserId` (`UserId`)，这就说明UserId在testtable1和testtable2表中都是唯一的，如果把testtable2的数据批量导入到testtable1，如果执行下面【导入1】的SQL，就会出现1062的错误，导致整个过程会回滚，没有达到导入差异数据的目的。

INSERT INTO testtable1(UserId,UserName,UserType)
SELECT UserId,UserName,UserType FROM testtable2;

u4_unique

（Figure4：Duplicate entry '101' for key 'IX_UserId'）

MySQL提供一个关键字：IGNORE，这个关键字判断每条记录是否存在，是否违反饿了表中的唯一索引，如果存在就不插入，而不存在的记录就会插入。

-- 导入2
INSERT IGNORE INTO testtable1(UserId,UserName,UserType)
SELECT UserId,UserName,UserType FROM testtable2;

所以执行完【导入2】，就会产生Figure5的结果，这已经达到了我们的目的了，但是你有没发现自增的ID值跳过了一些值，这是因为我们之前执行【导入1】失败造成的，虽然我们的事务回滚了，但是自增ID会出现断层。在SQL Server中也会有这样的问题。扩展阅读：简单实用SQL脚本Part：查找SQL Server 自增ID值不连续记录

u5_效果

（Figure5：IGNORE效果）

(二) 导入并覆盖重复数据，REPLACE INTO 的使用

1. 把testtable1和testtable2分别回滚到Figure2和Figure3的状态（使用TRUNCATE TABLE命名再执行Insert语句），这个时候再执行下面的SQL，看有什么效果：

-- 导入3
REPLACE INTO testtable1(UserId,UserName)
SELECT UserId,UserName FROM testtable2;

u6_rep

（Figure6：REPLACE效果）

从上图Figure6中，我们可以看到：UserId为101的记录发生了改变，不单UserName修改了，而且UserType也变为NULL了。

所以，如果导入中发现了重复的，先删除再插入，如果记录有多个字段，在插入的时候如果有的字段没有赋值，那么新插入的记录这些字段为空（新插入记录的UserType都为NULL）。

需要注意的是，当你replace的时候，如果被插入的表如果没有指定列，会用NULL表示，而不是这个表原来的内容。如果插入的内容列和被插入的表列一样，则不会出现NULL。

2. 如果我们表结构UserType字段不允许为空，而且没有默认值的情况，执行【导入3】会发生什么事情呢？

u7_not null

（Figure7：返回警告信息）

u8_0

（Figure8：UserType被设置为0）

通过Figure7和Figure8，我们知道数据记录还是插入了，只是返回Field 'UserType' doesn't have a default value的警告，插入记录的UserType字段都被设置为0（'UserType' 为int数据类型）。

3. 如果我们希望导入的时候一起更新UserType字段的值，这自然很简单了，使用下面的SQL脚本就可以解决：

-- 导入4
REPLACE INTO testtable1(UserId,UserName,UserType)
SELECT UserId,UserName,UserType FROM testtable2;

u9_rep

（Figure9：一起更新UserType）

(三) 导入保留重复数据未指定字段，INSERT INTO ON DUPLICATE KEY UPDATE 的使用

把testtable1和testtable2分别回滚到Figure2和Figure3的状态（使用TRUNCATE TABLE命名再执行Insert语句），这个时候再执行下面的SQL，看有什么效果：

-- 导入5
INSERT INTO testtable1(UserId,UserName)
SELECT UserId,UserName FROM testtable2
ON DUPLICATE KEY UPDATE
testtable1.UserName = testtable2.UserName;

u10_update

（Figure10：保留UserType值）

对比Figure2、Figure3与Figure10，UserId为101的记录：更新了UserName的值，保留了UserType的值；但是由于【导入5】中没有指定UserType，所以新插入记录的UserType是为NULL的。

-- 导入6
INSERT INTO testtable1(UserId,UserName,UserType)
SELECT UserId,UserName,UserType FROM testtable2
ON DUPLICATE KEY UPDATE
testtable1.UserName = testtable2.UserName;

u11_update

（Figure11：保留UserType值）

对比Figure2、Figure3与Figure11，只插入testtable2表的UserId,UserName字段，但是保留testtable1表的UserType字段。如果发现有重复的记录，做更新操作；在原有记录基础上，更新指定字段内容，其它字段内容保留。

(四) 总结

当在一个UNIQUE键上插入包含重复值的记录时，默认的insert会报1062错误，MYSQL可以通过以上三种不同的方式和你的业务逻辑进行处理。

三、参考文献

MYSQL插入处理重复键值的几种方法

bitsCN.com

Statement

The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn

Related Article

Adding Users to MySQL: The Complete Tutorial

Adding Users to MySQL: The Complete TutorialMay 12, 2025 am 12:14 AM

Mastering the method of adding MySQL users is crucial for database administrators and developers because it ensures the security and access control of the database. 1) Create a new user using the CREATEUSER command, 2) Assign permissions through the GRANT command, 3) Use FLUSHPRIVILEGES to ensure permissions take effect, 4) Regularly audit and clean user accounts to maintain performance and security.

Mastering MySQL String Data Types: VARCHAR vs. TEXT vs. CHAR

Mastering MySQL String Data Types: VARCHAR vs. TEXT vs. CHARMay 12, 2025 am 12:12 AM

ChooseCHARforfixed-lengthdata,VARCHARforvariable-lengthdata,andTEXTforlargetextfields.1)CHARisefficientforconsistent-lengthdatalikecodes.2)VARCHARsuitsvariable-lengthdatalikenames,balancingflexibilityandperformance.3)TEXTisidealforlargetextslikeartic

MySQL: String Data Types and Indexing: Best Practices

MySQL: String Data Types and Indexing: Best PracticesMay 12, 2025 am 12:11 AM

Best practices for handling string data types and indexes in MySQL include: 1) Selecting the appropriate string type, such as CHAR for fixed length, VARCHAR for variable length, and TEXT for large text; 2) Be cautious in indexing, avoid over-indexing, and create indexes for common queries; 3) Use prefix indexes and full-text indexes to optimize long string searches; 4) Regularly monitor and optimize indexes to keep indexes small and efficient. Through these methods, we can balance read and write performance and improve database efficiency.

MySQL: How to Add a User Remotely

MySQL: How to Add a User RemotelyMay 12, 2025 am 12:10 AM

ToaddauserremotelytoMySQL,followthesesteps:1)ConnecttoMySQLasroot,2)Createanewuserwithremoteaccess,3)Grantnecessaryprivileges,and4)Flushprivileges.BecautiousofsecurityrisksbylimitingprivilegesandaccesstospecificIPs,ensuringstrongpasswords,andmonitori

The Ultimate Guide to MySQL String Data Types: Efficient Data Storage

The Ultimate Guide to MySQL String Data Types: Efficient Data StorageMay 12, 2025 am 12:05 AM

TostorestringsefficientlyinMySQL,choosetherightdatatypebasedonyourneeds:1)UseCHARforfixed-lengthstringslikecountrycodes.2)UseVARCHARforvariable-lengthstringslikenames.3)UseTEXTforlong-formtextcontent.4)UseBLOBforbinarydatalikeimages.Considerstorageov

MySQL BLOB vs. TEXT: Choosing the Right Data Type for Large Objects

MySQL BLOB vs. TEXT: Choosing the Right Data Type for Large ObjectsMay 11, 2025 am 12:13 AM

When selecting MySQL's BLOB and TEXT data types, BLOB is suitable for storing binary data, and TEXT is suitable for storing text data. 1) BLOB is suitable for binary data such as pictures and audio, 2) TEXT is suitable for text data such as articles and comments. When choosing, data properties and performance optimization must be considered.

MySQL: Should I use root user for my product?

MySQL: Should I use root user for my product?May 11, 2025 am 12:11 AM

No,youshouldnotusetherootuserinMySQLforyourproduct.Instead,createspecificuserswithlimitedprivilegestoenhancesecurityandperformance:1)Createanewuserwithastrongpassword,2)Grantonlynecessarypermissionstothisuser,3)Regularlyreviewandupdateuserpermissions

MySQL String Data Types Explained: Choosing the Right Type for Your Data

MySQL String Data Types Explained: Choosing the Right Type for Your DataMay 11, 2025 am 12:10 AM

MySQLstringdatatypesshouldbechosenbasedondatacharacteristicsandusecases:1)UseCHARforfixed-lengthstringslikecountrycodes.2)UseVARCHARforvariable-lengthstringslikenames.3)UseBINARYorVARBINARYforbinarydatalikecryptographickeys.4)UseBLOBorTEXTforlargeuns

See all articles

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

Video Face Swap

Video Face Swap

Swap faces in any video effortlessly with our completely free AI face swap tool!

Hot Article

Roblox: Grow A Garden - Complete Mutation Guide

3 weeks agoByDDD

Roblox: Bubble Gum Simulator Infinity - How To Get And Use Royal Keys

3 weeks agoBy尊渡假赌尊渡假赌尊渡假赌

How to fix KB5055612 fails to install in Windows 10?

3 weeks agoByDDD

Nordhold: Fusion System, Explained

3 weeks agoBy尊渡假赌尊渡假赌尊渡假赌

Mandragora: Whispers Of The Witch Tree - How To Unlock The Grappling Hook

3 weeks agoBy尊渡假赌尊渡假赌尊渡假赌

Hot Tools

SecLists

SecLists

SecLists is the ultimate security tester's companion. It is a collection of various types of lists that are frequently used during security assessments, all in one place. SecLists helps make security testing more efficient and productive by conveniently providing all the lists a security tester might need. List types include usernames, passwords, URLs, fuzzing payloads, sensitive data patterns, web shells, and more. The tester can simply pull this repository onto a new test machine and he will have access to every type of list he needs.

ZendStudio 13.5.1 Mac

ZendStudio 13.5.1 Mac

Powerful PHP integrated development environment

MantisBT

MantisBT

Mantis is an easy-to-deploy web-based defect tracking tool designed to aid in product defect tracking. It requires PHP, MySQL and a web server. Check out our demo and hosting services.

MinGW - Minimalist GNU for Windows

MinGW - Minimalist GNU for Windows

This project is in the process of being migrated to osdn.net/projects/mingw, you can continue to follow us there. MinGW: A native Windows port of the GNU Compiler Collection (GCC), freely distributable import libraries and header files for building native Windows applications; includes extensions to the MSVC runtime to support C99 functionality. All MinGW software can run on 64-bit Windows platforms.

SublimeText3 Linux new version

SublimeText3 Linux new version

SublimeText3 Linux latest version

Hot Topics

1666

14

CakePHP Tutorial

1425

52

Laravel Tutorial

1327

25

1273

29

1252

24