search
HomeTechnology peripheralsAINew 'AI scientists” combine theory and data to discover scientific equations

The goal of scientists is to discover meaningful formulas that accurately describe experimental data. Mathematical models of natural phenomena can be created manually based on domain knowledge, or they can be created automatically from large data sets using machine learning algorithms. The academic community has studied the problem of merging related prior knowledge and related function models, and believes that finding a model that is consistent with prior knowledge of general logical axioms is an unsolved problem.

Researchers from the IBM research team and Samsung AI team developed a method "AI-Descartes" that combines logical reasoning with symbolic regression to extract data from axiomatic knowledge and experimental data. in principle derivation of models of natural phenomena.

The research is titled "Combining data and theory for derivable scientific discovery with AI-Descartes" and was published on April 12, 2023 in "Nature Communications》.

New AI scientists” combine theory and data to discover scientific equations

Artificial neural networks (NN) and statistical regression are often used to automatically discover patterns and relationships in data. NN returns a "black box" model, where the underlying functions are typically used only for prediction. In standard regression, the functional form is predetermined, so model discovery amounts to parameter fitting. In symbolic regression (SR), the functional form is not predetermined but consists of operators from a given list (e.g., , -, ×, and ÷) and is calculated from the data.

SR models are generally more "interpretable" than NN models and require less data. Therefore, to discover natural laws symbolically from experimental data, SR may be more effective than NN or fixed-form regression; the integration of NN and SR has been the subject of recent research in neurosymbolic AI. A major challenge in SR is identifying scientifically meaningful models from the many models that fit the data. Scientists define a meaningful function as one that balances accuracy and complexity. However, many such expressions exist for a given data set, and not all of them are consistent with known background theory.

An alternative approach is to start with a known background theory, but there are currently no practical inference tools that can generate theorems consistent with experimental data from a known set of axioms. Automatic Theorem Provers (ATP) are the most widely used reasoning tools that can prove conjectures for a given logical theory. Computational complexity is a major challenge for ATP; for some types of logic, proving conjectures is undecidable.

Additionally, deriving models from logical theories using formal reasoning tools is especially difficult when arithmetic and calculus operators are involved. Machine learning techniques have been used to improve the performance of ATP, for example, by using reinforcement learning to guide the search process.

Derivable models must not only be empirically accurate, but they should also be predictive and insightful.

Researchers from the IBM Research Team and Samsung AI Team attempted to obtain such a model by combining a novel mathematical optimization-based SR method with an inference system. This resulted in an end-to-end discovery system "AI-Descartes" that extracts formulas from data via SR and then provides a proof of the formula's derivability from a set of axioms, or provides a proof of inconsistency. When a model is provably not derivable, the researchers propose new measures that indicate how close the formula is to a derivable formula, and use their inference system to calculate the values ​​of these measures.

New AI scientists” combine theory and data to discover scientific equations

Illustration: System overview. (Source: paper)

In early work combining machine learning with inference, scientists used logic-based descriptions to constrain the output of GAN neural architectures that generated images. There are also teams that combine machine learning tools and inference engines to search for functional forms that satisfy pre-specified constraints. This is to augment the initial data set with new points, thus improving the efficiency of the learning method and the accuracy of the final model. Some teams also leverage prior knowledge to create additional data points. However, these studies only considered constraints on the functional form to be learned and did not include general background theoretical axioms (logical constraints describing other laws and unmeasured variables involved in the phenomenon).

Cristina Cornelio, lead author of the paper and a research scientist at Samsung AI, said AI-Descartes offers some advantages over other systems, but its most distinguishing feature is that it logical reasoning ability. If there are multiple candidate equations that fit the data well, the system identifies which equation best fits the background scientific theory. The ability to reason also sets the system apart from "generative AI" programs like ChatGPT, which have limited logic capabilities in large language models and sometimes mess with basic math.

"In our work, we are combining first-principles methods with the more common data-driven methods of the machine learning era, which have been used by scientists for centuries. "This combination allows us to leverage both approaches and create more accurate and meaningful models for a wide range of applications."

The name AI-Descartes is a tribute to the 17th-century mathematician and philosopher René Descartes, who believed that the natural world could be described by a few basic physical laws and that logical inferences played a key role in scientific discoveries. .

New AI scientists” combine theory and data to discover scientific equations

Illustration: Explanation of the scientific method for system implementation. (Source: Paper)

Researchers from this team have demonstrated that combining logical reasoning with symbolic regression is of great value in obtaining meaningful symbolic models of physical phenomena. ; because they are consistent with background theory and generalize well to domains significantly larger than experimental data. The combination of regression and inference produces better models than either SR or logical inference alone.

Improvement or replacement of individual system components and the introduction of new modules, such as abductive inference or experimental design will expand the functionality of the entire system. Deeper integration of inference and regression can help synthesize data-driven and first-principles-based models and lead to a revolution in the scientific discovery process. Discovering models that are consistent with prior knowledge will accelerate scientific discovery and transcend existing discovery paradigms.

The team used models to deduce Kepler's third law of planetary motion, Einstein's relativistic time dilation law, and Langmuir's adsorption theory; the research shows that when logical reasoning is used to When distinguishing candidate formulas with similar errors on the data, the model can discover dominant patterns from a small number of data points.

New AI scientists” combine theory and data to discover scientific equations

Illustration: Visualization of related sets and their distances. (Source: paper)

# "In this work, we need human experts to write down in a formal, computer-readable way what the axioms of the background theory are, and if If humans miss any of them or get any of them wrong, the system won't work," said Tyler Josephson, assistant professor of chemistry, biochemistry and environmental engineering at UMBC. "In the future, we also hope to automate this part of the job so we can Explore more fields of science and engineering."

Ultimately, the team hopes their AI-Descartes can inspire a productive new scientific approach just like real scientists. "One of the most exciting aspects of our work is the potential for significant advances in scientific research," Cornelio said.

Paper link: https://www.nature.com/articles/s41467-023-37236-y

Related reports: https://techxplore.com/news/2023-04-ai-scientist-combines-theory-scientific.html

The above is the detailed content of New 'AI scientists” combine theory and data to discover scientific equations. For more information, please follow other related articles on the PHP Chinese website!

Statement
This article is reproduced at:51CTO.COM. If there is any infringement, please contact admin@php.cn delete
ai合并图层的快捷键是什么ai合并图层的快捷键是什么Jan 07, 2021 am 10:59 AM

ai合并图层的快捷键是“Ctrl+Shift+E”,它的作用是把目前所有处在显示状态的图层合并,在隐藏状态的图层则不作变动。也可以选中要合并的图层,在菜单栏中依次点击“窗口”-“路径查找器”,点击“合并”按钮。

ai橡皮擦擦不掉东西怎么办ai橡皮擦擦不掉东西怎么办Jan 13, 2021 am 10:23 AM

ai橡皮擦擦不掉东西是因为AI是矢量图软件,用橡皮擦不能擦位图的,其解决办法就是用蒙板工具以及钢笔勾好路径再建立蒙板即可实现擦掉东西。

谷歌超强AI超算碾压英伟达A100!TPU v4性能提升10倍,细节首次公开谷歌超强AI超算碾压英伟达A100!TPU v4性能提升10倍,细节首次公开Apr 07, 2023 pm 02:54 PM

虽然谷歌早在2020年,就在自家的数据中心上部署了当时最强的AI芯片——TPU v4。但直到今年的4月4日,谷歌才首次公布了这台AI超算的技术细节。论文地址:https://arxiv.org/abs/2304.01433相比于TPU v3,TPU v4的性能要高出2.1倍,而在整合4096个芯片之后,超算的性能更是提升了10倍。另外,谷歌还声称,自家芯片要比英伟达A100更快、更节能。与A100对打,速度快1.7倍论文中,谷歌表示,对于规模相当的系统,TPU v4可以提供比英伟达A100强1.

ai可以转成psd格式吗ai可以转成psd格式吗Feb 22, 2023 pm 05:56 PM

ai可以转成psd格式。转换方法:1、打开Adobe Illustrator软件,依次点击顶部菜单栏的“文件”-“打开”,选择所需的ai文件;2、点击右侧功能面板中的“图层”,点击三杠图标,在弹出的选项中选择“释放到图层(顺序)”;3、依次点击顶部菜单栏的“文件”-“导出”-“导出为”;4、在弹出的“导出”对话框中,将“保存类型”设置为“PSD格式”,点击“导出”即可;

GPT-4的研究路径没有前途?Yann LeCun给自回归判了死刑GPT-4的研究路径没有前途?Yann LeCun给自回归判了死刑Apr 04, 2023 am 11:55 AM

Yann LeCun 这个观点的确有些大胆。 「从现在起 5 年内,没有哪个头脑正常的人会使用自回归模型。」最近,图灵奖得主 Yann LeCun 给一场辩论做了个特别的开场。而他口中的自回归,正是当前爆红的 GPT 家族模型所依赖的学习范式。当然,被 Yann LeCun 指出问题的不只是自回归模型。在他看来,当前整个的机器学习领域都面临巨大挑战。这场辩论的主题为「Do large language models need sensory grounding for meaning and u

ai顶部属性栏不见了怎么办ai顶部属性栏不见了怎么办Feb 22, 2023 pm 05:27 PM

ai顶部属性栏不见了的解决办法:1、开启Ai新建画布,进入绘图页面;2、在Ai顶部菜单栏中点击“窗口”;3、在系统弹出的窗口菜单页面中点击“控制”,然后开启“控制”窗口即可显示出属性栏。

强化学习再登Nature封面,自动驾驶安全验证新范式大幅减少测试里程强化学习再登Nature封面,自动驾驶安全验证新范式大幅减少测试里程Mar 31, 2023 pm 10:38 PM

引入密集强化学习,用 AI 验证 AI。 自动驾驶汽车 (AV) 技术的快速发展,使得我们正处于交通革命的风口浪尖,其规模是自一个世纪前汽车问世以来从未见过的。自动驾驶技术具有显着提高交通安全性、机动性和可持续性的潜力,因此引起了工业界、政府机构、专业组织和学术机构的共同关注。过去 20 年里,自动驾驶汽车的发展取得了长足的进步,尤其是随着深度学习的出现更是如此。到 2015 年,开始有公司宣布他们将在 2020 之前量产 AV。不过到目前为止,并且没有 level 4 级别的 AV 可以在市场

ai移动不了东西了怎么办ai移动不了东西了怎么办Mar 07, 2023 am 10:03 AM

ai移动不了东西的解决办法:1、打开ai软件,打开空白文档;2、选择矩形工具,在文档中绘制矩形;3、点击选择工具,移动文档中的矩形;4、点击图层按钮,弹出图层面板对话框,解锁图层;5、点击选择工具,移动矩形即可。

See all articles

Hot AI Tools

Undresser.AI Undress

Undresser.AI Undress

AI-powered app for creating realistic nude photos

AI Clothes Remover

AI Clothes Remover

Online AI tool for removing clothes from photos.

Undress AI Tool

Undress AI Tool

Undress images for free

Clothoff.io

Clothoff.io

AI clothes remover

AI Hentai Generator

AI Hentai Generator

Generate AI Hentai for free.

Hot Tools

mPDF

mPDF

mPDF is a PHP library that can generate PDF files from UTF-8 encoded HTML. The original author, Ian Back, wrote mPDF to output PDF files "on the fly" from his website and handle different languages. It is slower than original scripts like HTML2FPDF and produces larger files when using Unicode fonts, but supports CSS styles etc. and has a lot of enhancements. Supports almost all languages, including RTL (Arabic and Hebrew) and CJK (Chinese, Japanese and Korean). Supports nested block-level elements (such as P, DIV),

MantisBT

MantisBT

Mantis is an easy-to-deploy web-based defect tracking tool designed to aid in product defect tracking. It requires PHP, MySQL and a web server. Check out our demo and hosting services.

SAP NetWeaver Server Adapter for Eclipse

SAP NetWeaver Server Adapter for Eclipse

Integrate Eclipse with SAP NetWeaver application server.

Atom editor mac version download

Atom editor mac version download

The most popular open source editor

MinGW - Minimalist GNU for Windows

MinGW - Minimalist GNU for Windows

This project is in the process of being migrated to osdn.net/projects/mingw, you can continue to follow us there. MinGW: A native Windows port of the GNU Compiler Collection (GCC), freely distributable import libraries and header files for building native Windows applications; includes extensions to the MSVC runtime to support C99 functionality. All MinGW software can run on 64-bit Windows platforms.