Language Models for Multimessenger Astronomy-Reference-Cited by-同舟云学术

Language Models for Multimessenger Astronomy

Published:2023-05-01 Issue:3 Volume:11 Page:63
ISSN:2075-4434
Container-title:Galaxies
language:en
Short-container-title:Galaxies

Author:

Sotnikov Vladimir¹^ORCID,Chaikova Anastasiia²^ORCID

Affiliation:

1. JetBrains and Astroparticle Physics Lab, JetBrains Research, Paphos 8015, Cyprus

2. School of Computer Science & Engineering, Constructor University, 28759 Bremen, Germany

Abstract

With the increasing reliance of astronomy on multi-instrument and multi-messenger observations for detecting transient phenomena, communication among astronomers has become more critical. Apart from automatic prompt follow-up observations, short reports, e.g., GCN circulars and ATels, provide essential human-written interpretations and discussions of observations. These reports lack a defined format, unlike machine-readable messages, making it challenging to associate phenomena with specific objects or coordinates in the sky. This paper examines the use of large language models (LLMs)—machine learning models with billions of trainable parameters or more that are trained on text—such as InstructGPT-3 and open-source Flan-T5-XXL for extracting information from astronomical reports. The study investigates the zero-shot and few-shot learning capabilities of LLMs and demonstrates various techniques to improve the accuracy of predictions. The study shows the importance of careful prompt engineering while working with LLMs, as demonstrated through edge case examples. The study’s findings have significant implications for the development of data-driven applications for astrophysical text analysis.

Publisher

MDPI AG

Subject

Astronomy and Astrophysics

Link

https://www.mdpi.com/2075-4434/11/3/63/pdf

Reference32 articles.

1. The Astronomer’s Telegram (ATel) (2023, February 28). Available online: https://www.astronomerstelegram.org.

2. GCN: The Gamma-ray Coordinates Network (2023, February 28). Available online: https://gcn.nasa.gov/.

3. Amazon Mechanical Turk (2023, February 28). Available online: https://www.mturk.com/.

4. Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P.J. (2019). Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer. arXiv.

5. Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C.L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., and Ray, A. (2022). Training language models to follow instructions with human feedback. arXiv.

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Astronomical Knowledge Entity Extraction in Astrophysics Journal Articles via Large Language Models;Research in Astronomy and Astrophysics;2024-05-24