The article considers the problem of quality control of scientific plots on the example of published plots characterizing continuum radiation absorption, properties of weakly bound molecular complexes and absorption cross sections used in atmospheric chemistry. The tasks of systematization of graphical resources are formulated, the functionality of the GrafOnto information system containing graphical resources is described, statistical samples characterizing, in particular, the quality of graphical resources in the collection are presented. The analysis of the quality of citing plots and proximity estimates for pairwise comparisons of all plots in the collection is presented as well as the applied ontology of graphical resources.
— Graphical resources on the continuum absorption of water vapor and its mixtures published in 2011–2020 are described. Summary tables are presented that characterize the main parameters of the absorption coefficients and transmission functions in different spectral intervals, the temperature dependence of the absorption coefficient, and the equilibrium constant of the water dimer formation reaction. The features of the study of the continuum absorption in published works in this time interval are noted. In a concise form, the results of the assessment of the quality of the cited graphs are presented, which are described by four qualitative and quantitative attributes. Three citation procedures are characterized, two of which are computerized. A method for estimating the difference between the citing and cited plots and examples of pairs “citing and cited plots” with a quantitative assessment of the difference are presented.
Доклад посвящен описанию технологии построения программного решения, автоматизирующего рутинные процессы, связанные с развитием оригинальных моделей TSUNM3 и CTM. Рассматриваемая прикладная научная информационно-вычислительная веб-система Метео+ предоставляет компьютерную поддержку научных исследований, в первую очередь, профессионалам в вычислительной геофизике, занимающимся разработкой и усовершенствованием собственных вычислительных моделей. При этом формируемая коллекция результатов моделирования позволяет привлекать уже других специалистов для детального анализа. The report is devoted to describing the technology of building a software solution that automates routine processes related to the development of the original TSUNM3 and CTM models. The considered applied scientific information-computational web-system Meteo+ provides computer support of scientific research, first of all, to professionals in computational geophysics who are engaged in development and improvement of their own computational models. At the same time, the formed collection of modeling results allows to involve already other specialists for detailed analysis.
В докладе представлен метод количественной оценки качества цитирования научных графиков на примере собранной нами коллекции графических ресурсов в системе GrafOnto. В настоящее время коллекция содержит более 6000 примитивных графиков, большая часть которых входит в составные графики или составные рисунки. Среди этих графиков около полутора тысячи цитируемых. Для оценки качества цитирования нами предлагается количественная метрика среднего расхождения пары “цитируемый-цитирующий график”. The paper describes an approach to quantitative assessment of the quality of scientific plots citation, on the example of our collection of graphical resources in the GrafOnto system. Currently the collection contains more than 6000 primitive plots, most of which are included in composite plots or composite figures. Among these plots are about 1,500 citations. To evaluate the quality of citation we propose a quantitative metric of average divergence of the pair "original-cited plot".
The GrafOnto collection of water absorption plots published in 1981–2000 is briefly overviewed. This collection is hosted in the W@DIS information system (wadis.saga.iao.ru). Most of the ordinates of the plots belong to three groups of functions describing the wavenumber dependences of the absorption coefficient and the transmission function, the temperature dependence of the absorption coefficient, and the frequency dependence of the correction factor χ of the Lorentz profile. We present two ways of searching for plots in the collection: simple search by three publication attributes and attribute search by 14 properties, including the properties of a primitive plot, substance, function parameters, and properties of an information resource and a publication. This search is required for users to find desired primitive plots and is used when combining primitive plots into a composite plot accounting for user requirements.
For monitoring and short-term forecasting of the meteorological situation and atmospheric air quality near settlements, transport hubs and industrial facilities, the Meteo+ automated computing system is proposed, based on a mathematical model of the atmospheric boundary layer and an effective numerical method focused on the use of supercomputers. The mathematical model includes an impurity transport model with a reduced chemical mechanism and a non-hydrostatic mesoscale meteorological model with a modern moisture microphysics parameterization scheme. Examples of the successful application of the developed automated computing system in the numerical prediction of surface air quality deterioration in light winds and temperature inversions, as well as in the prediction of such dangerous weather phenomena as wind gusts are given.
The paper describes the infrastructure scientific informationcomputing web-system under development for the preparation, modeling, visualization and analysis of short-term weather forecast data to assess the quality of atmospheric surface air over a large urban agglomeration. The unified automated software complex Meteo+ under consideration is primarily designed for automation and management of infrastructure processes of collection, storage and analysis of the resulting numerical weather prediction modeling and atmospheric air quality assessment results. A common unified systematization of resources within the Meteo+ software package enables professionals in computational geophysics, who are engaged in collective development and improvement of their own computational models TSUNM3 and CTM, to focus on solving their application tasks, minimizing technical routine work on preparation, conducting and supporting calculations, and subsequent system storage, analysis and exchange of calculation results. In its final form, the unified automated Meteo+ software package described will offer: convenient online preparation and execution of computational models (TSUNM3, CTM and third-party WRF, CAMx); enhanced management and storage of computational results; and analysis and interpretation of data sets in tabular or graphical views.
В статье систематизированы и кратко описаны научные графики спектральных функций, относящиеся к континуальному поглощению углекислого газа и опубликованные в работах 1991-2000 гг. Выбранные графики доступны для исследователей в информационной системе W@DIS (wadis.saga.iao.ru). В рамках этой системы графики могут сравниваться между собой, а также с результатами расчетов или измерений, принадлежащих пользователям этой системы.
An approach to the formation of a description of the city's transport system is considered in order to identify the most polluted road spans with vehicle exhaust gases. The quantitative characteristics of pollution are determined in accordance with the standard Russian GOST R 56162-2019. When estimating emissions from road transport, information presented on the web in graphical form about traffic jam is used. In addition, the information on the number of cars, average speed and type of transport obtained by analyzing the video stream from web cameras by the YOLO neural network (version 4) is also used. The binding of pollution results to the transport system is formalized using the transport system ontology. The results can be used in dynamic models of pollution in urbanized areas.
. The paper discusses the prospects for the development of information systems associated with the Russian nodes that are part of the Virtual Atomic and Molecular Data Center. Atomic, ionic, and molecular spectral data are accumulated at these nodes. The key issues of the segment development are the semantization of tabular and graphical information resources and the personalization of data and its properties. This work lists the main information problems related to the three stages of the knowledge life cycle, reuse of accumulated data, and semantic search of spectral resources. It is shown how the issues of building researchers' own expert arrays based on the knowledge base of spectral information resources and physical quantities will be solved.
The report provides an overview of the results of work on the systematization of scientific plots characterizing the spectral functions related to the continuum absorption of carbon dioxide. Systematized plots are available for researchers in the W@DIS information system (wadis.saga.iao.ru) and can be used for comparison, both among themselves and with the results of calculations or measurements by users of this system. The work is a continuation of work on the systematization of graphic resources describing the continuum absorption by atmospheric molecules and contains dozens of little-known plots from publications of Soviet researchers in Russian.
Relations between the components of the data layer and application layer in the quantitative spectroscopy are schematically represented. The sequence of actions in the data quality analysis in the W@DIS information system is described. Methods for the spectral data quality analysis and assessment of trust in expert data sources and the results of their application in the W@DIS information system are presented. Along with common techniques for the primary data source quality analysis, techniques for decomposition of expert data sources and pairwise comparison of ordered data sets are reviewed. We also discuss the use of empirical data for filtering large data collections by an acceptable difference between identical energy levels in primary data sources and in an empirical data source. Two types of the user interface for viewing the results of spectral data analysis and assessment of trust in expert data are considered. Original methods for the data source quality analysis and assessment of trust in expert data are briefly described. These methods are used for finding discrepancies between different spectral data types (primary, expert, and empirical).
The report presents the tasks on graphical resources management thoroughly describing applied ontologies of GrafOnto research graphics collection used for solving problems of spectroscopy. Two groups of the tasks on graphical resources management are discussed in the paper. The first group is oriented on the quality analysis of graphical resources and the representation of these results as a graphical resources ontology. The second group of tasks is associated with the automatic class creation for the aforementioned ontology. The problems of ontology modularity and automatic classes' generation are being discussed. Examples of solving reduction problem as well as applied ontologies metrics are presented.
This paper presents an overview of the current status of the Virtual Atomic and Molecular Data Centre (VAMDC) e-infrastructure, including the current status of the VAMDC-connected (or to be connected) databases, updates on the latest technological development within the infrastructure and a presentation of some application tools that make use of the VAMDC e-infrastructure. We analyse the past 10 years of VAMDC development and operation, and assess their impact both on the field of atomic and molecular (A&M) physics itself and on heterogeneous data management in international cooperation. The highly sophisticated VAMDC infrastructure and the related databases developed over this long term make them a perfect resource of sustainable data for future applications in many fields of research. However, we also discuss the current limitations that prevent VAMDC from becoming the main publishing platform and the main source of A&M data for user communities, and present possible solutions under investigation by the consortium. Several user application examples are presented, illustrating the benefits of VAMDC in current research applications, which often need the A&M data from more than one database. Finally, we present our vision for the future of VAMDC.
An approach is presented which allows visual and quantitative trust assessment of composite multi-paper plots, where the results of the graphical representation of calculation or measurement results from different publications are compared. Using plots from the W@DIS IS collection, which characterize the continuum absorption, as an example, we show the procedure of construction of a composite multi-paper plot.
В докладе дано краткое описание созданной коллекции графиков, характеризующих спектральные процессы, и представлена оценка качества семантической коллекции графиков, характеризующих процессы поглощения в спектроскопии и результаты исследований потенциальных функций, описывающих слабо связанные комплексы. Ключевой проблемой, рассмотренной в докладе, является оценка качества цитируемых графиков. Выделены сложности, возникающие при оценке, и указаны не решенные задачи при корректной оценке качества научных графиков. This work briefly characterizes the spectral processes scientific plots collection and provides the quality assessment of the semantic collection of scientific plots. The collection contains data on spectroscopic absorption processes and the results of weakly bounded complexes potential functions research. The key issue of this report is the quality assessment of cited scientific plots. The difficulties with quality assessment are highlighted and unsolved problems with quality control for scientific plots are denoted.
В докладе обсуждаются методы анализа качества спектральных данных и оценка доверия экспертным данным и результаты их применения в ИС WDIS. Наряду с традиционными методами, используемыми при анализе качества первичных источников данных, обсуждается также метод декомпозиции экспертных источников данных, метод попарного сравнения упорядоченных массивов данных и использование эмпирических данных для фильтрации больших коллекций данных по величине допустимой разницы между уровнями энергии первичных источников данных и эмпирического источника. В докладе рассмотрены два типа интерфейсов для просмотра результатов анализа спектральных данных и оценки доверия экспертным данным. The report discusses methods for analyzing the quality of spectral data, trust assessment of expert data sources and the results of their application in the WDIS information system. Along with the traditional methods used in analyzing the quality of primary data sources, the method of decomposing expert data sources, the method of pairwise comparison of ordered data arrays, and the use of empirical data to filter large collections of data by the magnitude of the acceptable difference between the energy levels of primary data sources and an empirical source are also discussed. The report examines two types of interfaces for viewing the results of spectral data analysis and assessing trust in expert data.
We describe the current state and technical characteristics of Tropospheric Ozone Research (TOR) station, created 25 years ago to monitor atmospheric composition, basic meteorological variables, and other parameters. The multiyear observations showed that the air quality on the territory of Akademgorodok in Tomsk has been substantially degraded since the creation and development of the Special Economic Zone on its territory.
An approach to forming applied ontologies in subject domains in which data are presented in various forms of tables and scientific graphics is proposed. A description of the sources of data and information presented in this form is given. Using quantitative spectroscopy as an example, an approach to forming semantic annotations characterizing these sources is demonstrated. The major types of the sources are described. For scientific graphics, an approach to solving the problem of reducing and systematizing the graphic resources to search for plots in the subject domain is described. A partition into groups of functions used in the plots that are not interrelated with each other is constructed to define different spectral functions to be equivalent. The metrics of three applied ontologies of spectroscopy used in comparing data collections are briefly described.
Представлены описание и первые результаты разработки виртуальной вычислительно-информационной среды для анализа, оценки и прогноза последствий глобальных климатических изменений окружающей среды и климата в выбранном регионе. Созданная среда основана на информационных ресурсах трехслойной архитектуры. Intermediate results of the project aimed at improving methods of the detailed analysis, assessment and prediction of global climate change impact on the regional environment and climate are presented. New reliable interactive tools for in-depth statistical analysis and studying climate change impact obtained in this project will provide specialists, professionals, decision-makers and stakeholders with detailed climatic information. The project addresses the development of a topical virtual research environment (VRE) for the comprehensive study of ongoing and possible future climate change. It analyses the relevant subsequent effects. Such VRE will provide full topical informational required for studying regional economic, political and social consequences of the global climate change. The ultimate goal of this work is a design of hardware and software prototype supporting the topical virtual research environment for climate and environmental monitoring and analysis of the impact of climate change on socioeconomic processes on both local and regional scales. This VRE will integrate both already known and new archives of climate data sets with software realizations of traditional and advanced methods for statistical analysis of big spatial data sets. VRE prototype will provide scientists, decision-makers and stakeholders the access to processing resources and services for interactive analysis of geographically distributed spatial data through a web browser. It will present the results of the analysis using geoinformation technologies and ensure the systematization of spatial data and associated climate information. Also, the work describes an ontological approach to this systematization, which makes it possible to compare the semantics of meteorological and climatic parameters used in different collections and applied problems.