search
HomeDatabaseMysql TutorialSolution to the problem of failure to insert emoji expressions into MySQL

Emoji expressions are often encountered in our daily development, but recently I encountered a problem when inserting emoji expressions into mysql. I finally solved it by searching for relevant information, so I will share the process of solving this problem. This article mainly I will introduce to you the solution to the problem of MySQL failure to insert emoji expressions. Friends in need can refer to it.

Preface

I used to think that UTF-8 was a universal solution to character set problems until I encountered this problem recently. Recently, I was working on a crawler for Sina Weibo. When saving, I found that as long as the emoji expression is maintained, the following exception will be thrown:


Incorrect string value: '\xF0\x90\x8D\x83\xF0\x90...'

As we all know UTF-8 It is 3 bytes, which already includes most of the fonts we see every day. But 3 bytes are far from enough to accommodate all the text, so there is utf8mb4. utf8mb4 is a superset of utf8, occupying 4 characters. section, backward compatible with utf8. The emoji expressions we use every day are 4 bytes.

So when we insert data into the utf8 data table, it will report Incorrect string value This error.

It’s easy to find the solution through Google. The specific solution is as follows:

1. Modify the data The character set of the table is utf8mb4

. This is very simple. You can find a lot of modification statements online, but it is recommended to rebuild the table and use mysqldump -uusername -ppassword database_name table_name > table.sql Back up the corresponding data table and modify the character set of the table creation statement to utf8mb4, then mysql -uusername -ppassword database_name Reimport sql You can complete the character set modification operation.<br>

2. The MySQL database version must be 5.5.3 or above

Network All the articles above indicate that MySQL 5.5.3 or above is required to support utf8mb4. However, the database version I used is 5.5.18, which can still solve the problem in the end, so students should not rush to the operation and maintenance brother to upgrade the database first. Try to see if you can solve the problem by yourself.

3. Modify the database configuration file /etc/my.cnf and restart the mysql service

Mainly modify the default character set of the database, as well as the connection and query character set. [Mysql supports emoji upgrade encoding to UTF8MB4][1] This article has detailed setting methods, [In-depth Mysql character set settings ][2] This article has the function of each character set set in it. You can learn more about it.

##4. Upgrade MySQL Connector to 5.1.21 and above

Of all the above operations, the most critical is step 3, modifying the database configuration file, which probably



[client]
# 客户端来源数据的默认字符集
default-character-set = utf8mb4
[mysqld]
# 服务端默认字符集
character-set-server=utf8mb4
# 连接层默认字符集
collation-server=utf8mb4_unicode_ci
[mysql]
# 数据库默认字符集
default-character-set = utf8mb4

These configurations specify the character sets used by the pipelines through which data passes from the client to the server. Problems with each pipeline may cause insertion failure or garbled characters.


But many times, online databases cannot modify database files casually, so our operation and maintenance classmates decisively rejected my request to modify the database configuration file (T_T)


So I can only Solved it with code. The first step is to start with the character set specified when connecting to JDBC.



jdbc:mysql://localhost:3306/ding?characterEncoding=UTF-8

Mainly change UTF-8 to utf8mb4 for The Java Style Charset string should be able to solve the problem, right?


But unfortunately, Java JDBC does not have a character set for utf8mb4. When using UTF-8, it can be compatible with urf8mb4 and automatically Convert character set.


For example, to use 4-byte UTF-8 character sets with Connector/J, configure the MySQL server with character_set_server=utf8mb4, and leave characterEncoding out of the Connector/J connection string. Connector/J will then autodetect the UTF-8 setting. – [MySQL:Using Character Sets and Unicode][3]


Later, I did some popular science, and in every query request, you can Explicitly specify the character set used, use

set names utf8mb4 You can specify the character set of this link as utf8mb4, but this setting will become invalid after each connection is released.

The current solution is to explicitly call and execute

set names utf8mb4 when you need to insert utf8mb4, such as:


jdbcTemplate.execute("set names utf8mb4");
jdbcTempalte.execute("...");

It should be noted that when we use the ORM framework, due to performance optimization reasons, the framework will delay submission. Unless the transaction ends or the user actively calls forced submission, the

set names utf8mb4 responsible for execution will still not take effect. .

Here I am using myBatis, taking MessageDao as an example



// MessageDao
public interface MessageDao {
 @Update("set names utf8mb4")
 public void setCharsetToUtf8mb4();
 @Insert("insert into tb_message ......")
 public void insert(Message msg);
}
// test code
SqlSession sqlSession = sqlSessioFactory.openSession();
messageDao = sqlSession.getMapper(MessageDao.class);
messageDao.setCharsetToUtf8mb4();
// 强制提交
sqlSession.commit();
messageDao.insert(message);

At this point, the problem is solved...


Hey, it would be great if things could go so smoothly. In the project, the mybatis instance is managed by Spring, which means I can't get the sqlSession, which means I can't force the submission. And Because of the limitations of the Spring transaction framework, it does not allow users to explicitly call forced submission. I am still struggling with this problem.


There are two solutions:

  • Using AOP, when it is possible to insert 4-byte UTF8 characters, the prefix method executes set names utf8mb4, but this solution cannot yet determine that the AOP method will be implemented by Spring. Transaction management, and in the pre-method, whether the link obtained is the same session as the connection object obtained next.

  • Study the creation method of Spring JDBC and write one The hook executes set names utf8mb4 every time it creates a new database connection, thus ensuring that the character set has been set for each link obtained.

The above is the detailed content of Solution to the problem of failure to insert emoji expressions into MySQL. For more information, please follow other related articles on the PHP Chinese website!

Statement
The content of this article is voluntarily contributed by netizens, and the copyright belongs to the original author. This site does not assume corresponding legal responsibility. If you find any content suspected of plagiarism or infringement, please contact admin@php.cn
图文详解mysql架构原理图文详解mysql架构原理May 17, 2022 pm 05:54 PM

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了关于架构原理的相关内容,MySQL Server架构自顶向下大致可以分网络连接层、服务层、存储引擎层和系统文件层,下面一起来看一下,希望对大家有帮助。

mysql怎么替换换行符mysql怎么替换换行符Apr 18, 2022 pm 03:14 PM

在mysql中,可以利用char()和REPLACE()函数来替换换行符;REPLACE()函数可以用新字符串替换列中的换行符,而换行符可使用“char(13)”来表示,语法为“replace(字段名,char(13),'新字符串') ”。

mysql怎么去掉第一个字符mysql怎么去掉第一个字符May 19, 2022 am 10:21 AM

方法:1、利用right函数,语法为“update 表名 set 指定字段 = right(指定字段, length(指定字段)-1)...”;2、利用substring函数,语法为“select substring(指定字段,2)..”。

mysql的msi与zip版本有什么区别mysql的msi与zip版本有什么区别May 16, 2022 pm 04:33 PM

mysql的msi与zip版本的区别:1、zip包含的安装程序是一种主动安装,而msi包含的是被installer所用的安装文件以提交请求的方式安装;2、zip是一种数据压缩和文档存储的文件格式,msi是微软格式的安装包。

mysql怎么将varchar转换为int类型mysql怎么将varchar转换为int类型May 12, 2022 pm 04:51 PM

转换方法:1、利用cast函数,语法“select * from 表名 order by cast(字段名 as SIGNED)”;2、利用“select * from 表名 order by CONVERT(字段名,SIGNED)”语句。

MySQL复制技术之异步复制和半同步复制MySQL复制技术之异步复制和半同步复制Apr 25, 2022 pm 07:21 PM

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了关于MySQL复制技术的相关问题,包括了异步复制、半同步复制等等内容,下面一起来看一下,希望对大家有帮助。

带你把MySQL索引吃透了带你把MySQL索引吃透了Apr 22, 2022 am 11:48 AM

本篇文章给大家带来了关于mysql的相关知识,其中主要介绍了mysql高级篇的一些问题,包括了索引是什么、索引底层实现等等问题,下面一起来看一下,希望对大家有帮助。

mysql怎么判断是否是数字类型mysql怎么判断是否是数字类型May 16, 2022 am 10:09 AM

在mysql中,可以利用REGEXP运算符判断数据是否是数字类型,语法为“String REGEXP '[^0-9.]'”;该运算符是正则表达式的缩写,若数据字符中含有数字时,返回的结果是true,反之返回的结果是false。

See all articles

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

AI Hentai Generator

AI Hentai Generator

Generate AI Hentai for free.

Hot Article

Repo: How To Revive Teammates
1 months agoBy尊渡假赌尊渡假赌尊渡假赌
R.E.P.O. Energy Crystals Explained and What They Do (Yellow Crystal)
2 weeks agoBy尊渡假赌尊渡假赌尊渡假赌
Hello Kitty Island Adventure: How To Get Giant Seeds
1 months agoBy尊渡假赌尊渡假赌尊渡假赌

Hot Tools

Dreamweaver Mac version

Dreamweaver Mac version

Visual web development tools

VSCode Windows 64-bit Download

VSCode Windows 64-bit Download

A free and powerful IDE editor launched by Microsoft

MinGW - Minimalist GNU for Windows

MinGW - Minimalist GNU for Windows

This project is in the process of being migrated to osdn.net/projects/mingw, you can continue to follow us there. MinGW: A native Windows port of the GNU Compiler Collection (GCC), freely distributable import libraries and header files for building native Windows applications; includes extensions to the MSVC runtime to support C99 functionality. All MinGW software can run on 64-bit Windows platforms.

PhpStorm Mac version

PhpStorm Mac version

The latest (2018.2.1) professional PHP integrated development tool

SAP NetWeaver Server Adapter for Eclipse

SAP NetWeaver Server Adapter for Eclipse

Integrate Eclipse with SAP NetWeaver application server.