<?xml-model href='http://www.tei-c.org/release/xml/tei/custom/schema/relaxng/tei_all.rng' schematypens='http://relaxng.org/ns/structure/1.0'?><TEI xmlns="http://www.tei-c.org/ns/1.0">
	<teiHeader>
		<fileDesc>
			<titleStmt><title level='a'>LEGO – II. A 3mm molecular line study covering 100pc of one of the most actively star-forming portions within the Milky Way disc</title></titleStmt>
			<publicationStmt>
				<publisher></publisher>
				<date>09/01/2020</date>
			</publicationStmt>
			<sourceDesc>
				<bibl> 
					<idno type="par_id">10187239</idno>
					<idno type="doi">10.1093/mnras/staa1814</idno>
					<title level='j'>Monthly Notices of the Royal Astronomical Society</title>
<idno>0035-8711</idno>
<biblScope unit="volume">497</biblScope>
<biblScope unit="issue">2</biblScope>					

					<author>A T Barnes</author><author>J Kauffmann</author><author>F Bigiel</author><author>N Brinkmann</author><author>D Colombo</author><author>A E Guzmán</author><author>W J Kim</author><author>L Szűcs</author><author>V Wakelam</author><author>S Aalto</author><author>T Albertsson</author><author>N J Evans</author><author>S C Glover</author><author>P F Goldsmith</author><author>C Kramer</author><author>K Menten</author><author>Y Nishimura</author><author>S Viti</author><author>Y Watanabe</author><author>A Weiss</author><author>M Wienen</author><author>H Wiesemeyer</author><author>F Wyrowski</author>
				</bibl>
			</sourceDesc>
		</fileDesc>
		<profileDesc>
			<abstract><ab><![CDATA[ABSTRACT            The current generation of (sub)mm-telescopes has allowed molecular line emission to become a major tool for studying the physical, kinematic, and chemical properties of extragalactic systems, yet exploiting these observations requires a detailed understanding of where emission lines originate within the Milky Way. In this paper, we present 60 arcsec (∼3pc) resolution observations of many 3 mm band molecular lines across a large map of the W49 massive star-forming region (∼100 pc×100pc at 11kpc), which were taken as part of the ‘LEGO’ IRAM-30m large project. We find that the spatial extent or brightness of the molecular line transitions are not well correlated with their critical densities, highlighting abundance and optical depth must be considered when estimating line emission characteristics. We explore how the total emission and emission efficiency (i.e. line brightness per H2 column density) of the line emission vary as a function of molecular hydrogen column density and dust temperature. We find that there is not a single region of this parameter space responsible for the brightest and most efficiently emitting gas for all species. For example, we find that the HCN transition shows high emission efficiency at high column density (1022cm−2) and moderate temperatures (35K), whilst e.g. N2H+ emits most efficiently towards lower temperatures (1022cm−2; &lt;20K). We determine $X_{\mathrm{CO} (1-0)} \sim 0.3 \times 10^{20} \, \mathrm{cm^{-2}\, (K\, km\, s^{-1})^{-1}}$, and $\alpha _{\mathrm{HCN} (1-0)} \sim 30\, \mathrm{M_\odot \, (K\, km\, s^{-1}\, pc^2)^{-1}}$, which both differ significantly from the commonly adopted values. In all, these results suggest caution should be taken when interpreting molecular line emission.]]></ab></abstract>
		</profileDesc>
	</teiHeader>
	<text><body xmlns="http://www.tei-c.org/ns/1.0" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:xlink="http://www.w3.org/1999/xlink">
<div xmlns="http://www.tei-c.org/ns/1.0"><p>particular, transitions with high critical densities are regularly used in the literature to selectively study the relatively denser gas within molecular clouds. For example, it is typically assumed that the HCN (J = 1-0) molecular line transition is a particularly good tracer of 'dense gas' (number density of H 2 above 10 4 cm -3 ), owing to its high critical density for emission, and a chemical formation pathway that produces abundant HCN molecules within cold, dense gas (e.g. <ref type="bibr">Boger &amp; Sternberg 2005)</ref>. Under specific physical assumptions, the luminosity of HCN (J = 1-0), L HCN(1-0) can be used to estimate the dense molecular gas mass, M dg , within star-forming regions <ref type="bibr">(Gao &amp; Solomon 2004a,b;</ref><ref type="bibr">Wu et al. 2010)</ref>. A measure of the dense gas mass, along with the star formation rate (SFR), &#7744; * , has been used to constrain the models for star formation through, for example, the star formation efficiency (SFE) of dense gas ( &#7744; * /M dg ). However, with the advent of more sensitive radio telescopes that are capable of efficiently mapping large regions of the sky, it has been pointed out that the L HCN(1-0) &#8733; M dg relation may not hold across environments (e.g. <ref type="bibr">Kauffmann et al. 2017;</ref><ref type="bibr">Pety et al. 2017;</ref><ref type="bibr">Shimajiri et al. 2017;</ref><ref type="bibr">Evans et al. 2020;</ref><ref type="bibr">Nguyen-Luong et al. 2020)</ref>. The results from these authors suggest that the underlying assumption that molecular line emission originates solely from gas around and above the critical density is fundamentally flawed, Figure <ref type="figure">1</ref>. The region around W49 within the plane of the Milky Way. The upper panel shows a three-colour image across the Galactic plane, and a zoom-in on the mapped region of W49 indicated by the heavy black line (see Section 2). Shown are the Spitzer 8 &#956;m in blue, Spitzer 24 &#956;m in green, and Herschel 250 &#956;m in red <ref type="bibr">(Carey et al. 2009;</ref><ref type="bibr">Churchwell et al. 2009;</ref><ref type="bibr">Molinari et al. 2010)</ref>. Overlaid as red contours is the 250 &#956;m map in levels of <ref type="bibr">100,</ref><ref type="bibr">200,</ref><ref type="bibr">400,</ref><ref type="bibr">1000,</ref><ref type="bibr">2000</ref>, 10000 MJy sr -1 , which were chosen to best highlight the dense dust/gas distribution <ref type="bibr">(Molinari et al. 2010)</ref>. The labels on the zoom-in panel show the positions of the W49A star-forming region and the W49B supernova remnant. Shown in the main and zoom-in panels are scale bars of 100 and 10 pc at a distance of 11 kpc, respectively. The lower row, left-hand panel shows a map of the peak brightness temperature from the 12 CO molecular line, overlaid with black contours of 4, 8, 15, 22, 25 K. This 12 CO map highlights the extended nature of the molecular emission, which completely fills the mapped region. The angular beam size of 60 arcsec, or &#8764;3 pc at the distance of W49A (&#8764;11 kpc <ref type="bibr">;</ref><ref type="bibr">Gwinn, Moran &amp; Reid 1992;</ref><ref type="bibr">Zhang et al. 2013)</ref>, is shown as a black circle in the lower left of the lower left panel. The lower row, centre and right-hand panels show maps of the molecular hydrogen column density N H 2 (centre panel) and dust temperature T dust (right-hand panel) derived from Herschel observations. These maps have been smoothed to an angular resolution of 60 arcsec to match the resolution of the IRAM-30m observations. Black contours of N H 2 have been overlaid on the N H 2 map in levels of 1, 2.5, 5, 10, 50, 100 &#215; 10 21 cm -2 , and black contours of T dust have been overlaid on the T dust map in levels of <ref type="bibr">10, 15, 20, 25, and 30 K.</ref> despite this being made by many molecular line studies. Indeed, <ref type="bibr">Evans (1989)</ref> highlighted three decades ago that such an assumption breaks down out of the Rayleigh-Jeans limit in the sub-mm regime, where sub-thermal emission is the typical mode for emission before quickly becoming optically thick just above the critical density. If anything then, the opposite is true, and emission from molecular lines within the sub-mm regime traces gas with densities up to the critical density. Moreover, there could be a contribution to L Q from various additional excitation mechanisms (e.g. electron excitation; <ref type="bibr">Goldsmith &amp; Kauffmann 2017)</ref>.</p><p>Investigating the emission properties of these dense gas tracers -and many other commonly observed molecular line transitions -across a broad range of environments is, therefore, of current critical importance (e.g. <ref type="bibr">Nishimura et al. 2017;</ref><ref type="bibr">Watanabe et al. 2017)</ref>. This is of particular relevance now more than ever, as these dense molecular gas tracers are becoming routinely observable in other galaxies (e.g. <ref type="bibr">Bigiel et al. 2016)</ref>, where, even at the highest achievable spatial resolution for the closest disc galaxies, each line of sight contains a very broad range of environments (e.g. gas densities and temperatures; e.g. Jim&#233;nez-Donaire et al. 2019).</p><p>One of the primary aims of the Molecular Line Emission as a Tool for Galaxy Observations (LEGO) project is to further investigate the above issue by obtaining large maps of many commonly observed 3 mm band molecular lines, towards a selection of Milky Way star-forming regions that span a range of environmental properties (e.g. cloud density, star formation feedback, galactic environment, and metallicity). Building on the initial results from the Orion nebula presented by <ref type="bibr">Kauffmann et al. (2017)</ref>, we present in this work observations of the W49 massive star-forming region, which represents the most distant and, in terms of star formation rate, extreme source within the LEGO sample.</p><p>A three-colour image of the W49 region and its position within the galactic plane of the Milky Way is presented in the upper row of Fig. <ref type="figure">1</ref>. The bright blues (Spitzer 8 &#956;m emission) and greens (Spitzer 24 &#956;m emission) within this image highlight the regions of ongoing star formation, whilst the reds (Herschel 250 &#956;m emission) show the cold dust <ref type="bibr">(Carey et al. 2009;</ref><ref type="bibr">Churchwell et al. 2009;</ref><ref type="bibr">Molinari et al. 2010)</ref>. The mapped region (indicated by the black box) is of particular interest for a study of how the gas emission properties vary with environment, thanks to the large range of physical regimes  <ref type="bibr">[14.4, 20.5, 38</ref>.9] K T min dust , T mean dust , T max dust <ref type="bibr">(Section 3.1 and Appendix C)</ref> present within such a relatively small projected area on the sky. Table <ref type="table">1</ref> summarizes several properties of interest for the mapped region.</p><p>The W49A region highlighted in Fig. <ref type="figure">1</ref> is a very well-studied actively star-forming region that contains high-mass stars and associated H II regions (total stellar mass of 10<ref type="foot">foot_3</ref>-5 M ; e.g. <ref type="bibr">Homeier &amp; Alves 2005)</ref>. Thanks to these H II regions, it is one of the brightest known radio sources outside of the Galactic Centre (e.g. <ref type="bibr">Westerhout 1958</ref>). Therefore, W49A has been the subject of many radio continuum and recombination line investigations that have tried to characterize the embedded stellar objects and ionized gas properties <ref type="bibr">(Anantharamaiah 1985;</ref><ref type="bibr">De Pree, Mehringer &amp; Goss 1997;</ref><ref type="bibr">De Pree et al. 2000;</ref><ref type="bibr">Rugel et al. 2019)</ref>. Moreover, and of particular interest to the study presented here, the central part of this region has been the subject of several mm-wavelength spectroscopic observations, which find a wealth of molecular line emission <ref type="bibr">(Roberts et al. 2011;</ref><ref type="bibr">Nagy et al. 2012;</ref><ref type="bibr">Galv&#225;n-Madrid et al. 2013;</ref><ref type="bibr">Nagy et al. 2015)</ref>. Parallax distance measurements using H 2 O maser emission towards W49A place it at a distance of 11.11 +0.79  -0.69 kpc, locating it in a distant section of the Perseus arm near the solar circle in the first Galactic quadrant <ref type="bibr">(Gwinn et al. 1992;</ref><ref type="bibr">Zhang et al. 2013)</ref>. The W49B supernova remnant is also highlighted in Fig. <ref type="figure">1</ref>, and stands out as a bright feature in 24 &#956;m emission (i.e. green in the upper right panel).</p><p>The paper is organized as follows. Section 2 describes the IRAM-30m observations, the procedure for creating the integrated intensity maps, and the Herschel derived molecular hydrogen column density and dust temperature maps. Section 3 presents the results of the comparison of integrated intensities of the various molecular lines to the column density. Section 4 presents an analysis of where the majority of the emission is emitted within our mapped region, created by applying regional and dust temperature masks. Section 5 places these in the context of our understanding of molecular line emission. In Section 6, we summarize the main results of this work.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2">O B S E RVAT I O N S</head></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.1">Observations with the IRAM 30 m telescope</head></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.1.1">Data reduction</head><p>The observations presented here were taken as part of the Molecular Line Emission as a Tool for Galaxy Observations (LEGO) survey ). The full observational information of this survey, which covers &#8764;25 regions throughout the Galactic disc, will be summarized in a future publication <ref type="bibr">(Kauffmann et al., in preparation)</ref>. This survey was conducted using the Eight Mixer Receiver (EMIR; <ref type="bibr">Carter et al. 2012</ref>) on the 30 m telescope of the Instituto de Radioastronom&#237;a Milim&#233;trica (IRAM) on Pico Veleta, Spain. Data in the horizontal and vertical polarizations were acquired with the Fast Fourier Transform Spectrometer (FTS) operating in a wide-band mode that provides a spectral resolution of 200 kHz (i.e. 0.6 km s -1 &#8226; [&#957;/100 GHz] -1 in velocity). The EMIR receiver was operated in two set-ups, centred at local oscillator frequencies of 97.830 and 103.840 GHz. As a result, the observations cover in combination continuous frequency ranges of <ref type="bibr">86.10-99.89 and 101.78-115.57</ref> GHz. Relatively bright emission lines from a variety of astrophysically relevant molecules can be found at these frequencies. The transitions investigated by LEGO are listed in Table <ref type="table">2</ref>, which is created as described in Section 2.1.2. The telescope's intrinsic beam size<ref type="foot">foot_1</ref> is 24. 6 &#8226; (&#957;/100 GHz) -1 . The angular resolution of our data is lower, though, due to the processing described below.</p><p>The observations cover a total area of &#8764;30 arcmin &#215; 30 arcmin size around a central reference coordinate of 19 h 10 m 32. s 487, 9 &#8226; 05 36. 532. The On-The-Fly (OTF) mapping technique was used to create the mosaic image, using a dump time of 0.25 s and a scanning speed of 90 arcsec s -1 . The rows of the OTF map were spaced by 10 arcsec, corresponding to less than half the telescope's intrinsic resolution at the frequencies 115 GHz covered by our project. The OTF maps were taken by scanning along either the RA or Dec. axis of the equatorial coordinate system. Every point in the map was covered by at least two orthogonal scans, in order to reduce artefacts due to scan patterns.</p><p>The LEGO data reduction pipeline uses functionality from the Python programming language and from IRAM's GILDAS software suite (i.e. the CLASS package). <ref type="foot">3</ref> This approach permits use of Python's superior bookkeeping functionality to control the data flow, while it also takes advantage of the rich and well-tested data processing functionality within CLASS for low-level tasks. The pipeline, as well as the derived data products, will be released as part of future publications. In the very first processing step, observing logs are generated on the basis of header information of observed scans and data from IRAM's Telescope Access for Public Archive System (TAPAS). 4 These logs constitute the basis for all further data processing steps.</p><p>The pipeline processes the data separately for every emission line recorded in Table <ref type="table">2</ref>. Python scripts are used to identify, for a given emission line, the observations covering a particular source. CLASS is then employed to extract a velocity range of 350 km s -1 total width around a given spectral line (a larger range is extracted around Table <ref type="table">2</ref>. Information on the selected observed molecular lines, ordered by increasing rest frequency. Section 2.1.2 describes how the lines recorded in this table were selected, and how the line characteristics recorded here were obtained. Columns 1-10 show the name of each molecule, the transition information (used to refer to each transition throughout), the frequency of the transition, the upper energy level of the transition, the Einstein spontaneous decay coefficient, the collisional deexcitation rate coefficients at a kinetic temperature of 20 K, and the critical and effective densities for emission. a Transitions that are not available within the LAMDA data base have the corresponding information blanked. Additional information on the molecular line data base used within this work can be found in Table <ref type="table">A1</ref>. The full, machine-readable version of this table can be obtained from the supplementary online material.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Species</head><p>Transition</p><p>H </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Note. a</head><p>The two-level approximation of the critical density has been calculated by accounting for only the single downward collisional rate coefficient from the initial to final energy level (C 2lvl ul and n 2lvl crit ). The full critical density calculation accounts for all the possible downward transition collisional rate coefficients to and from the final energy level (C ul and n crit ). The effective excitation densities (n eff ) are taken from <ref type="bibr">Shirley (2015)</ref>, and have been defined by radiative transfer modelling as the density that results in a molecular line with an integrated intensity of 1 K km s -1 (also see <ref type="bibr">Evans 1999)</ref>.</p><p>transitions with substructure in order to contain all lines), to grid all data to a common set of velocity channels of 0.6 km s -1 width, and to subtract a spectral baseline of third order from the data. Velocities in the range -50 to + 80 km s -1 are ignored when fitting the baseline. This range is further expanded in case of emission lines with hyperfine structure, provided this line substructure is more compact than 20 km s -1 . The original velocity range of -50 to + 80 km s -1 is not modified for baseline fitting in cases where line substructure extends over a range &gt;20 km s -1 : we hypothesize that line substructure spread out over a large velocity range is less likely to bias baseline fits, and the absence of substantial artefacts in manually inspected processed data validates this approach. Future versions of the pipeline may include more sophisticated algorithms for automatic flagging of significant line emission during baseline fitting, and future publications will include a detailed assessment of the baseline quality.</p><p>The spectra delivered by the telescope's data processing system are calibrated in forward-beam brightness temperatures (the T * A -scale). We convert these to the main beam brightness scale using the relation</p><p>We adopt frequency-independent forward and main beam efficiencies of F eff = 0.95 and B eff = 0.8, respectively, consistent with the available calibration data. 2 The baselined data were eventually gridded into a data cube with pixels of 10 arcsec &#215; 10 arcsec size. This effectively also averages data from the two independent polarizations in every part of the map. This step uses a frequency-dependent Gaussian gridding kernel that increases the intrinsic resolution of the data (i.e. impact of frequency-dependent telescope beam and gridding) to 30 arcsec.</p><p>The angular resolution of our data is, however, further reduced due to the high scan speed and slow sampling pattern employed by our project. We pursue this fast mapping strategy because it allows us to cover a large area on the sky at the expense of somewhat reduced angular resolution. To be specific, along the scan direction, we obtain one data sample every 90 arcsec s -1 &#8226; 0.25 s = 22.5 arcsec. The impact of this sampling pattern on the angular resolution can, to first order, be gauged by modelling it by a Gaussian smoothing kernel of 22.5 arcsec width at half peak value that is extended along the scan direction. The resulting angular resolution in this modelalong the sampling direction of a single scan -would be ([30 arcsec] 2 + [22.5 arcsec] 2 ) 1/2 = 37.5 arcsec. The angular size of the beam would be smaller when averaging perpendicular scans. A future dataoriented LEGO paper will provide a more detailed discussion of this issue.</p><p>To allow direct comparison to the molecular hydrogen column density and dust temperature maps (see Section 3.1), we further smooth data cubes with a Gaussian kernel of 52 arcsec width at half peak value to achieve an angular resolution of 60 arcsec, and re-sample the data set on to the same 20 arcsec pixel size spatial grid. Further smoothing the data set also has the benefit of increasing the signal-to-noise of the lines across the mapped region, whilst somewhat mitigating the beam smearing effect caused by the fast mapping speed.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.1.2">Selection of emission lines</head><p>The LEGO pipeline generates data cubes spanning a few hundred channels for a selected set of emission lines. This approach is necessary because processing the full data set would be computationally too demanding. Table <ref type="table">2</ref> lists the transitions for which data cubes are produced. Additional details on quantum numbers and frequencies are presented in Table <ref type="table">A1</ref>. All of the lines are emitted by molecules, with the exception of the H41&#945; hydrogen recombination line.</p><p>The catalogue is systematically constructed using the following steps. At a minimum, LEGO should cover all emission lines that can be studied in nearby galaxies. We thus start our line catalogue by including all lines that <ref type="bibr">Watanabe et al. (2014)</ref> detect in M51 at a signal-to-noise ratio &gt;2. To this we add all transitions studied by <ref type="bibr">Pety et al. (2017)</ref> in Orion B, to be consistent with their work. We then include rare isotopologues of species where the abundant line is potentially optically thick. (This is not relevant in practice, since rare forms of HCN, HNC, and HCO + are already included due to previous selection steps.) We further include bright lines from selected species that intuitively seem interesting, such as long and simple carbon chains. In practice, these lines are taken from <ref type="bibr">Watanabe et al. (2015)</ref>. Finally, the lowest-lying hydrogen &#945; recombination line above a frequency of 86 GHz is included in the table.</p><p>We are dealing with a large number of transitions, and the nature of our experiments does not require a particularly high level of accuracy in transition properties. We, therefore, obtain transition parameters from catalogues that cover a large number of lines, but we do not make any effort to refine the accuracy of parameters on the basis of the latest literature. Rest frequencies are taken from the Lovas/NIST data base <ref type="bibr">(Lovas 2004)</ref>, as accessed via the Splatalogue 5 data base for astronomical spectroscopy. This catalogue is also used to characterize the substructure of a transition. Provided splitting is present in a given transition, the emission lines were split into groups of &lt;100 MHz width (i.e. &lt;300 km s -1 &#8226; [&#957;/100 GHz] -1 in velocity). The group's rest frequency, &#957; rest , is taken to be the frequency of the line with the highest 'intensity' in the <ref type="bibr">Lovas (2004)</ref> catalogue. The full quantum numbers of this transition are recorded in Table <ref type="table">A1</ref>. The lowest and highest frequencies of lines in a given group (i.e. &#957; min and &#957; max ) are also collected in Table <ref type="table">A1</ref>, for example, to guide baseline subtraction and to build data cubes that are wide enough to contain all transitions within a given line group. Table <ref type="table">2</ref> lists the quantum numbers common within a given line group, as well as &#957; rest .</p><p>The Leiden Atomic and Molecular Database <ref type="bibr">(Sch&#246;ier et al. 2005</ref>; LAMDA, accessed in November 2019) is used to obtain the upper energy levels, Einstein coefficients and downward collisional rates coefficients for the lines covered by LEGO. When calculating the collisional rates coefficients, we ignore line splitting, where present, and use LAMBDA data files collapsing transition substructure in those cases. The critical density for each of the transitions has been calculated using two approximations. The first is the simple two-level approximation, which only accounts for the 5 <ref type="url">https://www.cv.nrao.edu/php/splat/index.php</ref> downward collisional rate of the initial upper (u) to final lower (l) energy level such that n 2lvl crit = A ul /C ul . The second approximation accounts for all the possible upward and downward transition collisional rate coefficients from the initial energy level, which when neglecting background contribution and assuming optically thin conditions can be defined as n crit = A ul / u = k C uk (subscript uk represents a transition from the 'upper' energy level, u, to any allowed energy level, k, even where E k &gt; E u ). Collisional rates of upward transitions have been calculated using the formalism outlined by <ref type="bibr">Shirley (2015, equation 4)</ref>. In both the calculations, we use the collision rates determined at a kinetic temperature 20 K, assuming that H 2 is the dominant collisional partner, and assume that the emission is optically thin. These results are presented in Table <ref type="table">2</ref>.</p><p>Also given in Table <ref type="table">2</ref> are the effective excitation densities (n eff ) taken from <ref type="bibr">Shirley (2015)</ref>. These have been determined using the RADEX radiative transfer modelling code <ref type="bibr">(van der Tak et al. 2007)</ref>, and are defined as the density that results in a molecular line with an integrated intensity of 1 K km s -1 (also see <ref type="bibr">Evans 1999)</ref>. These estimates then account for radiative trapping effects that lower the critical density within the optically thick regime. In Table <ref type="table">2</ref>, we show n eff for a kinetic temperature 20 K and the reference column densities for each molecule from Shirley (2015, table 1).</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.1.3">Spectra, noise properties, and integrated intensity maps</head><p>A brief inspection of the cubes shows that they contain a varying degree of complexity in their position-position-velocity space structure. Fig. <ref type="figure">2</ref> shows the spectrum from each molecular transition averaged across the mapped region. Here, we can see that several of the lines contain multiple velocity components, which vary in spatial structure and peak intensity (also see channel maps presented in Fig. <ref type="figure">D1</ref>).</p><p>Highlighted at the top of Fig. <ref type="figure">2</ref> is the range of velocities (&#8764;-10 to 20 km s -1 ) observed towards W49A, and represented as a vertical dashed black line is the mean systemic velocity of W49A (&#8764;11 km s -1 ; e.g. <ref type="bibr">Nagy et al. 2012</ref><ref type="bibr">Nagy et al. , 2015))</ref>. Within this velocity range, we find asymmetric line profiles, and, in some cases, multiple narrowly separated (&lt;10 km s -1 ) velocity components. These line profiles are likely due to the unresolved internal structure and kinematics (e.g. outflows) within W49A, which is currently being heavily influenced by stellar feedback (e.g. see <ref type="bibr">Galv&#225;n-Madrid et al. 2013)</ref>. Also labelled in Fig. <ref type="figure">2</ref> are examples of several more broadly spaced (&gt;10 km s -1 ) velocity components from sources unrelated to W49A, which, given their velocities, are thought to arise from emission in intervening spiral arms (e.g. W49B at &#8764;60 km s -1 ).</p><p>There are a couple of additional potential causes for the multiple velocity components seen in averaged spectra. The first is hyperfine splitting of molecular transitions (also see Table <ref type="table">A1</ref>). Labelled in Fig. <ref type="figure">2</ref> are several examples for lines that have narrow (e.g. N 2 H + ) and broadly (e.g. CCH) spaced hyperfine transitions. The second is self-absorption within regions of high optical depth, which can result in the narrow (&lt;10 km s -1 ) spaced velocity components. This would be particularly relevant towards the centre of W49A, and for the more abundant molecules (e.g. CO; <ref type="bibr">Nagy et al. 2012;</ref><ref type="bibr">Galv&#225;n-Madrid et al. 2013)</ref>.</p><p>The final feature to note within Fig. <ref type="figure">2</ref>, is the blending of the CH 3 OH spectra. The CH 3 OH-E and CH 3 OH-A transitions shown here are very close in frequency (corresponding to a velocity Figure 2. Normalized spectra from each molecule averaged over the mapped region, for pixels with a integrated intensity &gt;3 &#963; W Q (see Fig. <ref type="figure">3</ref>). The peak brightness temperature of each spectrum is given on the left axis. The coloured shaded region of each spectrum shows &gt;0 K intensities. Highlighted are several features of interest within the spectra, such as the systemic velocity of W49A, hyperfine transitions, and line blending.</p><p>difference of &#8764;5 km s -1 ), and cannot be fully resolved towards W49A (also see Fig. <ref type="figure">3</ref>). <ref type="foot">6</ref> Further analysis of the CH 3 OH lines in this work is, therefore, limited to the significantly brighter A-type line, and all references to CH 3 OH (2-1) henceforth correspond to the CH 3 OH-A (2-1) line.</p><p>Despite identifying several velocity components within the spatially averaged spectra, when examining the cubes we find that, in general, there is a single velocity component that dominates the emission along a given line of sight. Such that the total emission at any position within a molecular line cube can typically be attributed to a single velocity range (see Fig. <ref type="figure">D1</ref>). It then appears the brightest sources within the mapped region are generally distinct in both velocity space and spatially. Throughout this work, we then make the simple assumption that at each position there is a single source that is responsible for the emission seen in molecular line emission. Additionally, the analysis presented in this work relies on a comparison to the total hydrogen column density and dust temperature maps, which also integrate emission from all molecular material along the line of sight (see Section 2.2). Hence, not limiting our analysis to a given velocity range allows us to make a consistent comparison across all the available data sets.</p><p>The aim of this work is to produce intensity maps integrated along all velocities that (a) trace the compact and bright emission, (b) recover the diffuse and extended emission, (c) have a constant noise profile where emission is not present. To do so, we follow a two-step masking procedure for the data cubes to reduce the noise within positions with significant line emission. Briefly, for this procedure, we firstly create a mask including only the voxels (velocity pixels) with significant CO emission, which we then apply to a given line cube. As the CO line shows significant emission within the majority of velocity channels where emission is found from the other lines, this masking procedure provides a convenient way of producing a generous integration velocity range at each line-of-sight such that both the noise is reduced in pixels containing emission and pixels without emission show a flat noise profile. In the case of lines with hyperfine structure that extends outside of the CO mask (e.g. CCH), we choose to expand the mask to also include voxels with significant emission from the line cube (i.e. producing a final mask where each position has either significant emission from the line hyperfine structure or CO emission).</p><p>We follow the method outlined by <ref type="bibr">Dame (2011)</ref> to produce the initial CO data cube mask. For this method, we firstly smooth the cubes by a factor of two spatially (i.e. to 120 arcsec), and a factor of 10 spectrally (i.e. to 6 km s -1 ). We then determine the rms value, &#963; rms , along each line of sight using two spectral windows, covering velocities of -150 to -20 km s -1 and 90-150 km s -1 . These velocity ranges were chosen for their lack of significant emission (see Fig. <ref type="figure">2</ref>). An initial smoothed CO cube mask is then produced to include all channels above a 5 &#963; rms threshold. If hyperfine structure is present within a given cube, but outside of this CO mask, then this smoothing procedure is applied to the cube. Significant emission from this smoothed cube is then included in the final mask for that cube only. These final smoothed masks are then applied to the full resolution cubes, which are used to create the integrated intensity maps (W Q = T Q,mb dv).  <ref type="table">3</ref>). The angular beam size of 60 arcsec, or &#8764;3 pc at the distance of W49A (&#8764;11 kpc; <ref type="bibr">Gwinn et al. 1992;</ref><ref type="bibr">Zhang et al. 2013)</ref>, is shown as a white circle in the lower left of the first panel.</p><p>Table <ref type="table">3</ref>. Observational properties across the mapped region (i.e. that covered with both vertical and horizontal on-the-fly scans). The columns show the molecule name, the average cube rms (in a 1 km s -1 channel), mean uncertainty of the integrated intensity, and mean value of the integrated intensity. The final column shows the area percentage, A cov , within the mapped region that has an integrated intensity above five times the uncertainty; i.e. W Q &gt; 5&#963; W Q . The table has been ordered by decreasing A cov value (cf. table 8 of <ref type="bibr">Pety et al. 2017)</ref>. The information given is for maps that have been smoothed to an angular resolution of 60 arcsec, and have a spectral resolution of 0.6 km s -1 . See Table <ref type="table">B1</ref> for additional statistical properties of the molecular line integrated intensity maps. The full, machine-readable version of this table can be obtained from the supplementary online material.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Line</head><p>&#963; rms (1 km s -1 ) The uncertainty on the integrated intensity at each pixel was calculated as,</p><p>where vres is the velocity resolution, and N ch is the number of unmasked channels used to create the moment maps. Mean values for the &#963; rms and &#963; W Q for each molecular line are shown in Table <ref type="table">3</ref>. The final moment maps that are used throughout this work are displayed in Fig. <ref type="figure">3</ref>.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="2.2">Dust column density and temperature maps</head><p>The molecular hydrogen column density and dust temperature maps, together with their uncertainties, were obtained by fitting modified blackbody models on a pixel-by-pixel basis to the HerschelSpace Observatory large program Hi-Gal data <ref type="bibr">(Molinari et al. 2011)</ref>. For that, we follow a similar procedure as described in <ref type="bibr">Guzm&#225;n et al. (2015)</ref>. We also subtract dust emission arising from diffuse gas in the Galactic plane attributable to clouds located in the foreground or background of W49 using the same method as in <ref type="bibr">Guzm&#225;n et al. (2015)</ref>. Here, however, the background is smoothed to a much larger angular scale of 660 arcsec (34 pc at a distance of 11 kpc), to approximately match the size of the mapped region. The subtracted maps balance preserving small-scale structures with an angular resolution equal to the longest Herschel wavelength map (32 arcsec), whilst removing large-scale variations due to the background (&gt;660 arcsec). A smoother background would leave too much diffuse emission not associated with W49 but to the Galactic plane, and a background with more structure on smaller scales would filter out emission associated with W49 itself. The column density and temperature maps produced using this method, however, suffer from minor small-scale artefacts, such as saturated values (e.g. towards the centre of W49A). To correct for these issues, we spatially smooth the maps using the ASTROPY.CONVOLUTION package, which accurately interpolates over bad or missing values within an image. We find that a Gaussian kernal of 52 arcsec provided the optimal tradeoff between final map resolution and a complete interpolation over missing values. The final molecular hydrogen column density and dust temperature have an angular resolution of 60 arcsec, and have been resampled on to a 20 arcsec pixel grid. These maps are presented within the lower row of Fig. <ref type="figure">1</ref>.</p><p>In Appendix C, we present an analysis of the column density and dust temperature uncertainties produced by random noise in the Herschel measurements. We find that mean fractional uncertainties across the mapped region studied in this work are N H 2 /N H 2 = &#177; 30 per cent and T dust /T dust = &#177;10 per cent (see Fig. <ref type="figure">C2</ref>).</p><p>It is useful to convert the H 2 column density map into units of visual extinction, A v , for comparison with other works (e.g. <ref type="bibr">Kauffmann et al. 2017)</ref>. To do so, we adopt the commonly used extinction relation</p><p>where A v is in units of mag, and N H 2 is in units of cm -2 <ref type="bibr">(Bohlin, Savage &amp; Drake 1978;</ref><ref type="bibr">Fitzpatrick 1999;</ref><ref type="bibr">Lacy et al. 2017)</ref>.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3">R E S U LT S</head><p>In this section, we present the results from the molecular hydrogen column density, dust temperature, and molecular line integrated intensity maps. For ease of comparison, all the data presented in this work have been smoothed to the same angular resolution of 60 arcsec, which corresponds to a spatial resolution of &#8764;3 pc at the distance of the W49A star-forming region (11 kpc; <ref type="bibr">Zhang et al. 2013)</ref>. It is worth noting that, given the low spatial resolution of these observations, we are most likely not fully resolving any structures within our source; e.g the central star-forming core of the W49A is contained within a single pixel (e.g. <ref type="bibr">Rugel et al. 2019)</ref>. Each pixel within our maps is expected, therefore, to contain a range of physical properties (e.g. densities and temperatures), and should be considered as parsec scale averages across a representative star-forming region of the Galactic Plane. The maps presented throughout this work contain all emission integrated along the line of sight. Similar to the effect of the low spatial resolution, this may cause averaging of physical properties within any given pixel.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3.1">Distribution of the molecular hydrogen and dust temperature</head><p>In Section 2, we discussed the H 2 column density uncertainty produced by the random noise in the Herschel maps. There is, however, a second systematic uncertainty produced by the unrelated fore-and back-ground clouds along the line of sight. This has already been somewhat accounted for by subtracting the diffuse emission from the individual Herschel wavelength maps, yet unrelated Galactic plane emission may be still present within the final column density map.</p><p>To account for the remaining foreground and background contamination, we implement a location-dependent minimum N H 2 (A v ) threshold for Galactic plane contamination. This threshold represents the Galactic background level on top of which the column density enhancements can be reliably attributed to the emission within the molecular line cubes. This threshold is defined, N thresh H 2 (A thresh v ), such that 90 per cent of the pixels within an area of the Galactic plane have column densities less than this threshold (i.e. N H 2 (90 per cent) &lt; N thresh H 2</p><p>). This analysis has been limited to the -0.4&lt;l &lt; 0.3 &#8226; , 43.5&lt; b &lt; 47.5 &#8226; region for the Galactic plane (see Fig. <ref type="figure">1</ref>). This is a large section of the Galactic plane directly adjacent to the LEGO mapped region, and hence should provide a representative background Galactic plane column density distribution for the same latitude range of the mapped region.</p><p>The upper panel of Fig. <ref type="figure">4</ref> shows the cumulative (pixel) distribution of molecular hydrogen column density (extinction on upper axis) across the mapped region and the Galactic plane. Highlighted as a horizontal dashed line is the 90 percentile for the Galactic plane, which corresponds to a column density N thresh H 2 = 2.5 &#215; 10 21 cm -2 (A thresh v = 2.7 mag). We find minimum-mean-maximum values of the column density within the mapped region of 0.1-2.5-104.2 &#215; 10 21 cm -2 , and dust temperatures of 14.4-20.6-38.8 K. Across the Galactic plane region, we find minimum-mean-maximum values of 0. <ref type="table">1-1.4-29.0 &#215; 10 21 cm -2 ,</ref> and<ref type="table">14.3-20.5-38.7 K</ref>. Whilst above the column density threshold within the mapped region we find 2.5-5.5-104.2 &#215; 10 21 cm -2 , and 14.7-21.8-38.8 K.</p><p>The central panel of Fig. <ref type="figure">4</ref> shows the dust temperature distribution as a function of column density for the mapped region and the Galactic plane. The overlaid contours highlight where the majority of data points lie within both distributions. Comparing these contours we see that both regions appear to be qualitatively similar, with the mapped region shifted higher by a factor of a few in column density. Moreover, below the column density threshold, there appears to be a shallow negative correlation between the column density and dust temperature for both regions. Such a negative correlation is typically expected for quiescent molecular gas, where the highest column densities have the lowest temperatures due to the lack of a heating source and efficient cooling mechanisms (i.e. via dust radiation). Given the aforementioned line-of-sight contamination at low column densities, it is, however, difficult to assess the significance of this relation.</p><p>Finally, the lower panel of Fig. <ref type="figure">4</ref> shows the dust temperature distribution as a function of column density for only the mapped region. This clearly shows that above the column density threshold there is an increase in the dust temperature with increasing column density, which originates from positions associated with the W49A massive star-forming region (cf. Fig. <ref type="figure">1</ref>). This is then indicative of an internal heating source associated with star formation (an increased interstellar radiation field; e.g. <ref type="bibr">Pety et al. 2017)</ref>.</p><p>We now compare to an independent measure of the molecular hydrogen column density. In doing so, we aim to test the fore-and back-ground contamination threshold of N thresh H 2 = 2.5 &#215; 10 21 cm -2 (A thresh v = 2.7 mag) in the Herschel molecular hydrogen column density map. To do so, we follow the method presented in <ref type="bibr">Heiderman et al. (2010, their equations 11-13)</ref> for converting the 13 CO(1-0) integrated intensity into a 13 CO column density. This method solves for the excitation temperature and optical depth of 13 CO(1-0) using both 12 CO(1-0) and 13 CO(1-0) peak intensities, by assuming LTE conditions, that both lines have equal excitation temperatures, and that the 12 CO(1-0) and 13 CO(1-0) are optically thick and thin, respectively. We then convert to a molecular hydrogen column density assuming a 13 CO abundance of 5 &#215; 10 -6 . Fig. <ref type="figure">5</ref> shows the molecular hydrogen column density determined from the 13 CO(1-0) observations, N H 2 ( 13 CO) , as a function of the molecular hydrogen column density determined from the Herschel observations, N H 2 . The overlaid solid red line shows the running median within logarithmically spaced N H 2 bins, and the diagonal black dashed line shows N H 2 ( 13 CO) = N H 2 . Comparing these two lines, we find a remarkably good agreement between N H 2 ( 13 CO) and N H 2 across around two orders of magnitude in molecular hydrogen column density. This indicates that the method for subtracting foreground and background emission within in the Herschel maps has correctly removed the Galactic cirrus emission across the map, leaving only the emission associated with the molecular hydrogen. Moreover, this agreement between N H 2 ( 13 CO) and N H 2 highlights that the same sources within each line of sight are recovered within both the integrated intensity and the molecular hydrogen column density maps. This then confirms our choice to integrate over all velocities within the molecular line cubes to produce the integrated intensity maps (Section 2.1.3).</p><p>We find that the N H 2 ( 13 CO) and N H 2 in Fig. <ref type="figure">5</ref> begin to deviate by more than the uncertainty on the Herschel column density at around N H 2 &lt; 10 21 cm -2 (see Section C). This value is significantly below N thresh H 2 = 2.5 &#215; 10 21 cm -2 , which highlights this as a reasonable threshold for the increased uncertainty &gt;30 per cent within the Herschel molecular hydrogen map. In summary, in this section, we have been able to differentiate the column density peaks within the mapped region from the systematically lower column density material associated with the background Galactic plane. We defined a column density threshold above which we expect that the measured column density and temperature values can be predominately attributed to a single source along any given line of sight, and, therefore, a more direct comparison can be made to the molecular line integrated intensity maps. This limit is highlighted on all figures within this work that make use of the column density within the variable plotted on the y-axis. Any values below the threshold should be assumed to have an associated uncertainty of &gt;30 per cent and, therefore, be taken with caution.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3.2">Distribution of the integrated intensities</head><p>The integrated intensity maps for each molecular transition are shown in Fig. <ref type="figure">3</ref> (as labelled). The logarithmic colour-scale for each panel has been adjusted to best display the features within each map. Signal-to-noise contours are overlaid at levels of 5 &#963; W Q , increasing up to the maximum &#963; W Q value within the mapped region (see Table <ref type="table">3</ref>).</p><p>We find that all the lines have peak integrated intensities towards the centre of W49A, at a position of 19 h 10 m 13 s , 9 &#8226; 06 16 . This corresponds to the position of a well-studied ring-like cluster of ultracompact H II regions, which is typically referred to as the 'Welch ring' (e.g. <ref type="bibr">Welch et al. 1987;</ref><ref type="bibr">Rugel et al. 2019)</ref>. This is in agreement with many extensive molecular line studies towards this central region, who find that both higher-J transitions and less abundant isotopologues of the molecule also peak towards this position (e.g. <ref type="bibr">Nagy et al. 2012</ref><ref type="bibr">Nagy et al. , 2015;;</ref><ref type="bibr">Galv&#225;n-Madrid et al. 2013)</ref>.</p><p>To assess the spatial extent of each line, we define its coverage as the percentage of pixels across the mapped region that have integrated intensities at least five times higher than their associated uncertainty values;</p><p>where A is the percentage of the total mapped area. We find that the CO lines are the most spatially extended across the mapped region, with both 12 CO and 13 CO covering 100 per cent and C 18 O covering around A cov &#8764; 35 per cent of the main mapped region (i.e. that covered with both vertical and horizontal onthe-fly scans). The next most extended lines are HCN, HCO + and CS, which each have coverages of between A cov &#8764; 20-30 per cent. These are followed by the HNC, CN, SO, CCH, and N 2 H + lines that cover around A cov &#8764; 10 per cent of the mapped region. Finally, the remaining 14 lines studied in this work cover only a few per cent of the mapped region. These coverage values, along with the minimum, mean, and maximum integrated intensity values measured within each mapped region are summarized in Table <ref type="table">3</ref>, which is ordered by decreasing coverage values.</p><p>Fig. <ref type="figure">6</ref> presents the mean integrated intensity and coverage as a function of critical density. Shown as triangles and circles are the critical densities determined using the two-level approximation (n 2lvl crit ) and when accounting for all the possible downward collisional rates (n crit ), respectively. For the lines given by <ref type="bibr">Shirley (2015)</ref>, we also plot the effective excitation density (n eff ) (see Table <ref type="table">3</ref>). This figure shows that beyond the 12 CO, 13 CO, and C 18 O lines there does not appear to be a clear correlation between the strength or coverage of an observed molecular transition and its critical or effective excitation density. This highlights that the simple approximations we have used to estimate the broad range of characteristic densities are not sufficient to explain the distribution of all the observed molecular lines. Such a result is not surprising, and could be easily explained by the lower abundance and, therefore, optical depth, of rarer molecules. This can be simply seen by studying the CO isotopologues in Fig. <ref type="figure">6</ref>, which show a stark decrease in mean brightness and coverage with increasing rareness despite having very similar critical densities (i.e. 12 CO, 13 CO, C 18 O, and C 17 O). This behaviour is discussed further in Sections 5.1 and 5.2.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="3.3">Molecular hydrogen column density versus integrated intensities</head><p>In Fig. <ref type="figure">7</ref>, we present scatter plots of the integrated intensities as a function of molecular hydrogen column density (or extinction as shown on the upper x-axis; Section 2.2). The black and grey circular points show the data with an integrated intensity value above and below 5&#963; W Q , respectively. We plot solid coloured lines that show the median values of the integrated intensity within equally spaced logarithmic column density (extinction) bins (shown at the top of each panel), where the shaded region shows the standard deviation range with each bin. To allow reliable detections of the integrated intensity down to low values of the column density, we stack the individual integrated intensity values within the logarithmic column density (extinction) bins. When doing so, the uncertainty on the median value, &#963; bin (med.), is defined as the propagated uncertainty from all the points contained within the bin, The triangles, circles, and squares show the critical density when using the two-level approximation (n 2lvl crit ) and accounting for all the possible collisional rates (n crit ), and the effective excitation density (n eff ), respectively (see Table <ref type="table">3</ref>). The coverage is defined as the percentage of pixels within the mapped region that have an integrated intensity higher than at least five times its associated uncertainty (W Q &gt; 5&#963; W Q ). The colours of each marker correspond to the molecular lines shown in the legend, and highlighted with labels in the figure are several molecular lines of interest. A linear y &#8733; x -1 relationship has been plotted for comparison. This plot highlights that using the critical density as a predictor for line brightness or coverage across a mapped region is not trivial. Rather, the abundance of a given molecule should be taken into account when predicting its emission characteristics.</p><p>where n is the number of points within the bin, and &#963; W Q (i) is the uncertainty associated with the ith value within the bin (see equation 1). The oversampling factor is defined as the ratio of the angular beam area to the pixel area, f os = 1.13 beam / pixel , and corrects for the number of independent measurements within each bin. Using the stacking method we find that even within bins that contain a significant fraction of points with &lt;5&#963; W Q (grey points), we can still extract a median integrated intensity value with a high level of significance. The result of this can be seen as the open coloured circles, which show the column density bins that have a median value above three times the associated uncertainty.</p><p>The first obvious trend from the binned lines in Fig. <ref type="figure">7</ref> is that on average there is a positive correlation between the integrated intensities and the H 2 column density. Indeed, we find that for each line the highest value of the integrated intensity is found at the position of the highest column density. This shows that, in general, the integrated intensities are relatively well correlated with the amount of material along the line of sight. However, there are significant deviations from a simple linear relationship.</p><p>To investigate these deviations from a simple linear relation, we normalize the integrated intensity by the H 2 column density. <ref type="bibr">Kauffmann et al. (2017)</ref> define this fraction h Q = W(Q)/N H 2 , which can be thought of as a proxy for the line emissivity, or the efficiency of an emitting transition per H 2 molecule. Fig. <ref type="figure">8</ref> shows the emission efficiency ratio as a function of column density (extinction; cf. fig. <ref type="figure">2</ref> of <ref type="bibr">Kauffmann et al. 2017)</ref>. As in Fig. <ref type="figure">7</ref>, we stack within logarithmically spaced column density bins to increase the significant detections at lower column densities. This is shown as the solid coloured lines and shaded regions, and the significant bins are shown with an open circle symbol. In Fig. <ref type="figure">8</ref>, the y-axis of each panel has been normalized such that the maximum value of the significant logarithmically spaced bins is equal to unity.</p><p>We find that the 12 CO, 13 CO, and C 18 O show a near-constant decrease in h Q with increasing column density, which is to be expected based on the overall sub-linear relations shown in Fig. <ref type="figure">7</ref>. Focusing on the significant median bins, we find that CS, CH 3 OH, HCO + , and N 2 H + show h Q profiles that quickly rises from low column densities to maximum emission efficiency at a column density of several 10 22 cm -2 , and slowly decrease towards higher column density values.</p><p>The profiles of HCN, HNC, CN, and CCH show more complex double-peaked profiles. These lines show an initially high value of h Q at the lowest values of the column density (&#8764;10 20-21 cm -2 ), that decreases with increasing column density to reach a minimum value at &#8764;10 22 cm -2 . The h Q then rises again to a maximum value of at column density around few 10 22 cm -2 . We highlight again here, however, that values of the column density below the threshold value of N thresh H 2 &#8764; 10 21 cm -2 suffer from significant fore-and background contamination (see Section 3.1). Therefore, these multiple peaked profiles for the emission efficiency should be taken with caution.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4">ANALYSIS</head></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.1">Fraction of the emission over various column densities</head><p>In an attempt to determine the typical extinction (column density) that is traced by a given molecular line transition, <ref type="bibr">Kauffmann et al. (2017)</ref> assess the cumulative fraction of the line intensity as a function of extinction (column density); W Q (N H 2 &lt; N H 2 ). Here we follow the same method and determine the cumulative distribution of the integrated intensity for each line. We note that, as we integrate all the intensity along the line of sight without isolating sources (e.g. W49A) in velocity, we do not have a single distance that could be reasonably applied to each pixel. Therefore, instead of producing the cumulative function of the luminosity as done in <ref type="bibr">Kauffmann et al. (2017)</ref>, we produce the cumulative function of the integrated intensity values across the mapped region. These results are displayed in Fig. <ref type="figure">9</ref>, where we split the observed molecular lines between four panels for plotting clarity.</p><p>Shown as shaded regions in Fig. <ref type="figure">9</ref> is the uncertainty on each cumulative distribution, which is determined as the deviation in the distributions when adding synthetic noise to both the molecular line and column density maps. To produce these noise curves, we first include synthetic noise in the form of a Gaussian distributions centred on zero with widths of &#963; W Q (Section 2.1.3) and N H 2 to each of the integrated intensity maps and the column density map, respectively. This was repeated 1000 times to produce a sample of noisy maps. The shaded curves in Fig. <ref type="figure">9</ref> show the 5, 25, 75, and 95 percentiles of the resultant ensemble of cumulative distributions.</p><p>Here the brightest lines (e.g. the CO lines) show a small uncertainty range (shaded region), as they are not significantly affected by the addition of &#963; W Q . To highlight the rigorous detection of these brighter lines (e.g. N 2 H + ), the inset panel shows the marginal increase in uncertainty with the addition of 10 &#963; W Q noise. The weaker lines (e.g. HC 3 N), on the other hand, show a larger variation due to the additional uncertainty, as the peak signal within these maps is closer to the &#963; W Q level (compare integrated intensity uncertainty values given in Table <ref type="table">3</ref> to integrated intensity statistics given in Table <ref type="table">B1</ref>). We also plot a dotted black line in Fig. <ref type="figure">9</ref> that shows the cumulative distribution of the pixels, and hence would represent the distribution of a line with a completely smooth integrated intensity map (i.e. a map containing pixels with only a single value). We plot the cumulative distribution of the column density values as a dashed black line in Fig. <ref type="figure">9</ref>. Using these, we can assess how much of the molecular line intensity is contained below a given column density; or in the case of W49, how concen-trated the emission is towards the column density peak in the W49A star-forming region. At any given column density, if the cumulative distribution profile resides between the smooth (dotted) and column density (dashed) profiles its remaining luminosity should be thought of as still being relatively evenly distributed across the mapped region (i.e. has a smoother integrated intensity map), whilst those to the right of the column density distribution should be thought of as having luminosities that are more strongly concentrated (i.e. has a more peaked integrated intensity map).</p><p>Fig. <ref type="figure">9</ref> shows that the cumulative distributions vary significantly among the different lines. We find that the CO lines are typically less concentrated than the column density (mass) distribution, whereas the remaining lines are typically more concentrated.</p><p>In Fig. <ref type="figure">10</ref>, we show the fraction of the total intensity, W frac Q , from each of the molecular lines within four column density (extinction) masks, which have been chosen for direct comparison with <ref type="bibr">Pety et al. (2017, their fig. 7</ref>). These authors suggest that within the Orion molecular cloud these ranges of column density (or extinction) can be used to represent diffuse (1 mag &lt; A V &lt; 2 mag), and translucent (2 mag &lt; A V &lt; 6 mag) gas, the environment of filaments (6 mag &lt; A V &lt; 15 mag), and dense gas (15 mag &lt; A V &lt; 222 mag). 7 We find that the CO lines have the majority of their intensity (W frac Q &#8764; 80 per cent) within the lowest two column density bins, which combined includes the majority of the mapped area (&#8764;70 per cent). The HCN, HCO + , HNC, CS, and CN lines show a stark decrease in the intensity within the lowest column density bin (W frac 12 CO &#8764; 30 per cent, W frac HCN &#8764; 10 per cent), and an increase within the highest column density bin (W frac 12 CO &#8764; 5 per cent, W frac HCN &#8764; 30 per cent). The N 2 H + shows an even higher fraction of emission towards the two highest bins, with W frac N 2 H + &#8764;40 per cent of the total emission within the highest column density bin that accounts for only around 1 per cent of the total mapped area.</p><p>These results are compared to the Orion B star-forming region <ref type="bibr">(Pety et al. 2017)</ref>, which is shown in the lower panel of Fig. <ref type="figure">10</ref> (also with the W49 results in grey). We see that there are several notable differences between the two regions. First, there appears to be a lower fraction of the intensity within the dense gas for the CO lines in W49 when compared to Orion. This is expected given that 7 The 1 mag &lt; A V &lt;2 mag and 2 mag &lt; A V &lt;6 mag extinction bins contain emission from below our extinction threshold of A thresh v = 2.7 mag (see Section 3.1). Fig. <ref type="figure">5</ref> shows that across 1&lt; A V &lt;2.7 mag N H 2 ( 13 CO) deviates from N H 2 by 30-40 per cent, which we suggest sets an upper limit to the uncertainty on W frac Q due to foreground and background contamination within this extinction range.</p><p>we observe a much larger spatial coverage within W49 (&#8764;10 4 pc 2 ) than the Orion observations (&#8764;10 pc 2 ), and hence recover a larger fraction of diffuse gas (relative to the dense gas). Secondly, we see that lines such as CN, HNC, HCO + , and HCN have a marginally higher fraction of the intensity within dense gas. Lastly, we observe a significant difference in the fraction of dense gas for the N 2 H + line, which has more than 80 per cent of the luminosity within the highdensity regime in Orion compared to around 40 per cent across the mapped region. <ref type="bibr">Watanabe et al. (2017), and</ref><ref type="bibr">Nishimura et al. (2017)</ref> conducted similar analyses of the 3 mm lines within the W51 and W3(OH) regions within the Galaxy. These authors also find results broadly in agreement with those found here. For example, <ref type="bibr">Nishimura et al. (2017)</ref> classify their observed distributions into the three types: (i) species concentrated just around the centre of the star-forming core in W3(OH), such as CH 3 OH, SO, and HC3N, (ii) species loosely concentrated around the centre, such as HCN, HCO + , HNC, and CS, and (iii) species extended widely, such as CCH, C 18 O, and 13 CO.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="4.2">Line emissivities as a function of both molecular hydrogen column density and dust temperature</head><p>So far in this work, we have investigated the variation of the emission from each molecular line as a function of the column density (extinction) and the critical or effective excitation density. However, as <ref type="bibr">Pety et al. (2017)</ref> showed within the Orion B region, the temperature is also a fundamental parameter, as it also is closely linked to both the excitation of molecular lines and the formation chemistry (i.e. abundances). In this section, we investigate the simultaneous effects of the molecular hydrogen column density and dust temperature on the total line emission and emission efficiency. We remind the reader that spatial distributions of both the molecular hydrogen column density and dust temperature are presented in Fig. <ref type="figure">1</ref>, and a comparison of the W49 and a representative portion of the Galactic plane is given in Fig. <ref type="figure">4</ref> (see Section 3.1). It is worth noting here, however, that the dust temperature may not be equal to the kinetic temperature of the gas. Without a probe of the latter within our observations (e.g. NH 3 or H 2 CO), we work under the assumption the dust and gas temperatures are coupled across the mapped region.</p><p>To conduct this analysis, we first take the integrated intensity (W Q ) and emission efficiency (h Q ) maps and identify the value of the column density and dust temperature for each position. We then create evenly spaced bins for the observed range of column density and dust temperature, into which we assign the W Q and h Q values based on their column densities and dust temperatures (i.e. similar to the binning process outlined in Section 3.3). Finally, we calculate the summed W Q and median h Q values within each column-density temperature bin, and calculate the uncertainty on this value using equation (1) (see Section 3.3). We note that both HNCO lines and the C 17 O and HN 13 C line have been omitted from this analysis due to their lack of significant column density and temperature bins.</p><p>Figs 11 and 12 show the parameter space spanned by the dust temperature and column density (extinction). The overlaid contours show the 1, 10, 20, 30 ... 80, and 90 per cent levels of the maximum summed W Q and median h Q for a given line across the significant bins within the parameter space (5&#963; bin ; see equation 2). These plots highlight the environmental properties which are producing the majority of the molecular line emission, and the conditions which are most efficiently producing this emission, respectively. Fig. <ref type="figure">11</ref> shows that many of the lines emit the majority of their emission from moderate temperatures (&#8764;20 K), and moderate column densities (&#8764;10 21-22 cm -2 ) following the peak in the number of pixels within the parameter space (compare to Fig. <ref type="figure">4</ref>). In addition, many of the lines also peak towards the highest column density and temperature region of the parameter space (&#8764;20 K; &#8764;10 23 cm -2 ). This region of the parameter space contains significantly fewer pixels, emphasizing that this peak is due to the concentration of bright pixels towards the centre of the W49A massive star-forming region.</p><p>Comparing how the total emission (Fig. <ref type="figure">11</ref>) and emission efficiency (Fig. <ref type="figure">12</ref>) are distributed across the observed parameter space, we find the column density and temperature ranges producing the most emission may not do so most efficiently. For example, although CO and 13 CO are producing most of the emission within the moderate density and temperature regime, they have a peak in emission efficiency at almost an order of magnitude lower column density. In addition, we find that many of the lines that showed two peaks in the emission distribution have a single peak in the emission efficiency distribution (e.g. HCO + , HCN, HNC, CS). Interestingly, several of the lines show comparable distributions for the emission and emission efficiency (N 2 H + , SO, CCH), highlighting that for these lines the gas properties producing the most emission are also doing so most efficiently.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5">D I S C U S S I O N</head><p>In this section, we discuss the results of this work in the context of our current understanding of the origins of molecular line emission from the interstellar medium. We note that an outline of the uncertainties present within our observations and the caveats that these imply to our interpretations is presented at the end of this section.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.1">Are 'dense gas' molecular line tracers preferentially tracing the 'dense' gas?</head><p>When studying the array of molecular line transitions observable within molecular clouds, the critical density of a given transition is typically used as a proxy of the density from which it is emitting. As such the different molecular line transitions can be used to probe the physical conditions of the various density regimes present throughout the cloud. Of particular interest are the properties of the 'dense gas' most closely associated with star formation. Much of our understanding of star formation is based on the study of these dense gas properties. For example, the star-forming potential of the molecular clouds is directly determined by the physical conditions (e.g. level of turbulence) of the dense gas, upon which the majority of the current analytic models for star formation are calibrated. Moreover, the mass determined from these dense gas tracers is used to derive empirical scaling relations to the level of star formation, which are widely used to calibrate simulations and estimate the star formation rate within extragalactic observations. The fundamental assumption typically adopted by molecular line studies is that the emission from any dense gas tracer, as inferred from its critical density, is (a) emitting from gas at that density (i.e. n &#8764; n crit ), and (b) this gas is directly linked to the star formation process (e.g. <ref type="bibr">Gao &amp; Solomon 2004b)</ref>. In this work, we address the first of these assumptions, and ask the question: are 'dense gas' molecular line tracers preferentially tracing the 'dense' gas?</p><p>In Fig. <ref type="figure">13</ref>, we present integrated intensity maps for a selection of the observed molecular lines, which have been ordered by increasing critical density (all integrated intensity maps are shown in Fig. <ref type="figure">3</ref>). Visual extinction, A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag)</p><p>Figure <ref type="figure">11</ref>. The distribution of integrated intensities as a function of the dust temperature and molecular hydrogen column density (extinction), where the integrated intensities have been binned along each axis to form a contour plot (cf. scatter plot shown in Fig. <ref type="figure">4</ref>). The contours show the 1, 10, 20, 30 ... 80, and 90 per cent levels of the total integrated intensity for a given line across the significant bins within the parameter space (5&#963; bin ; see equation 1). This plot highlights the properties of the region from which the majority of molecular line emission across the mapped region are originating.</p><p>This plot was made to be comparable to fig. <ref type="figure">1</ref> of <ref type="bibr">Kauffmann et al. (2017)</ref>, which presents the molecular line emission across the local Orion A star-forming region. Here, we can see broadly similar trends to what was found for Orion A. We find CO, 13 CO, have extended emission across all the mapped region, whilst 17 CO emits only towards the centre of W49A. Given that these molecular lines have comparable critical densities, this effect is likely due to the reduced abundance of the rarer isotopologues and insufficient integration time. Or, it could be that the rarer species are destroyed in the lower density material, reducing further their abundance (fractionation effects, e.g. <ref type="bibr">Langer et al. 1984;</ref><ref type="bibr">R&#246;llig &amp; Ossenkopf 2013)</ref>.</p><p>Moving beyond these CO lines in Fig. <ref type="figure">13</ref>, we find little correspondence between the extent of the molecular line emission and the increase in critical density. For example, we see comparable spatial morphologies of the CCH, CN, and HCN lines, despite them having critical densities differing by almost two orders of magnitude. Moreover, we find that only N 2 H + appears to also trace the column density peak at the east of the mapped region (towards W49B) at comparable brightness to the peak towards the centre of the W49A; potentially highlighting this as a region of differing excitation or chemistry than W49. These results for all the observed lines are summarized in Fig. <ref type="figure">6</ref>, from which we find that there is a significant scatter around the simple linear relation of the average integrated intensity across the mapped region and increasing critical density; several extreme outliers being C 18 O and CH 3 OH. This scatter is increased when plotting the coverage across the mapped region (i.e. positions with an integrated intensity above five times the uncertainty) as a function of the critical density. Together, Figs 6 and 13 then clearly show that the brightness and coverage of a molecular line transition are not completely governed by the critical density; e.g. excitation effects or chemical abundance variations may also be important factors. This is in agreement with the results from the Orion molecular cloud <ref type="bibr">(Kauffmann et al. 2017;</ref><ref type="bibr">Pety et al. 2017)</ref>. A v (mag) 10 -1 10 0 10 1 10 2</p><p>Visual extinction, A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag) 10 -1 10 0 10 1 10 2</p><p>A v (mag)</p><p>Figure <ref type="figure">12</ref>. The distribution of emission efficiencies as a function of the dust temperature and molecular hydrogen column density (extinction), where the emission efficiencies have been binned along each axis to form a contour plot (cf. scatter plot shown in Fig. <ref type="figure">4</ref>) The contours show the 1, 10, 20, 30 ... 80, and 90 per cent levels of the emission efficiency [h Q = W(Q)/N H 2 ] for a given line across the significant bins within the parameter space (5&#963; bin ; see equation 1). We highlight here values of the column density below N thresh H 2 = 2.5 &#215; 10 21 cm -2 (A thresh v = 2.7 mag) that suffer from increased (&gt;30 per cent) uncertainties due to foreground and background contamination (see Section 3.1). Values of the h Q below this threshold value should be taken with caution. This plot highlights the properties from which the molecular lines are most efficiently emitting across the mapped region are originating.</p><p>In Section 3.3, we investigated the emissivity of the molecular lines. When doing so, we identified large deviations from a simple linear relation for the integrated intensity of a molecular line and the column density. To investigate this, we normalize the integrated intensity by the column density, which we defined as the emission efficiency ratio [h Q = W(Q)/N H 2 ; <ref type="bibr">Kauffmann et al. 2017</ref>]. In Fig. <ref type="figure">14</ref>, we present the emission efficiency ratio as a function of the column density for a selection of the observed molecular lines, which have plotted on a single axis for comparison (comparable to Fig. <ref type="figure">8</ref>). This plot shows a variety of profiles for the different molecular lines with increasing column density. Also shown in the lower panel of Fig. <ref type="figure">14</ref> are the emission efficiency ratios determined within the Orion A star-forming region taken from <ref type="bibr">Kauffmann et al. (2017, their fig. 2)</ref>. We find a reasonable first-order agreement between W49 and Orion A. We see that 12 CO has a maximum emission efficiency ratio at the lowest column density. C 18 O, HCN, and N 2 H + show peak emission efficiency ratio values at increasingly higher column densities. There are, however, several second-order differences that should be mentioned. The first is that we do not observe a peaked emission efficiency ratio profile for 12 CO within W49, as is seen within Orion (comparing blue curves in Fig. <ref type="figure">14</ref>). Rather, we find a gradual decline of the 12 CO emission efficiency ratio with increasing column density. The second is that for Orion the HCN emission efficiency ratio peaks at a lower column density than N 2 H + , whereas in W49 we find these are comparable, with N 2 H + even potentially peaking at lower column density values. It is worth noting here that the spatial resolution of the W49 observations is &#8764;3 pc, whereas the beam of the Orion A observations is &#8764;0.05 pc. Therefore, it is encouraging that we are even broadly recovering the trends seen within Orion.</p><p>In Section 4.2, we built upon this analysis by assessing how the total intensity and emission efficiency ratio at a given column density show the column density and dust temperature determined from the Herschel observations. Overlaid on the upper left panel are contours N H 2 in levels of 1 and 2.5 &#215; 10 21 cm -2 , and in all panels levels of 5, 10, 50 &#215; 10 21 cm -2 . The filled contours for the molecular line emission are drawn at signal-to-noise ratios of 3, 5, 10, 30, 50, 100, 150, 200, and 300. All maps shown here have been smoothed to 60 arcsec (&#8764;3 pc at the source distance of 11 kpc), and the panels are ordered by increasing critical density. Labelled within each of the integrated intensity maps the peak value (see Table <ref type="table">B1</ref>).</p><p>varies as a function of dust temperature. In Fig. <ref type="figure">15</ref>, we present a simplified version of Figs 11 and 12. Here, we show the total intensity in the upper panel, and the emission efficiency ratio in the lower panel, as a function of column density and temperature for 12 CO, C 18 O, HCN, and N 2 H + on a single axis to allow comparison. To highlight the highest intensity peaks and most efficiently emitting gas we only show contours above a 50 per cent level.</p><p>We first see from Fig. <ref type="figure">15</ref> that CO, C 18 O, and HCN produce the majority of their total emission at moderate temperatures (&#8764;20 K), and moderate column densities (&#8764;10 21-22 cm -2 ). This region of the parameter space contains the most number of pixels (compare to the Fig. <ref type="figure">4</ref>), and hence shows that the majority of the emission from these lines is not strongly affected by smaller scale local intensity enhancements. The majority of N 2 H + emission across the mapped region, on the other hand, comes from higher column densities and is split between moderate and high temperatures. These results may then have implications for understanding molecular line emission from unresolved measurements at or above the size scale of the mapped region (&#8764;100 pc at 11 kpc).</p><p>Comparing the total intensities to efficiency ratios in the lower panel of Fig. <ref type="figure">15</ref>, however, we find very different distributions. We see that CO has an emissivity that peaks towards the lowest column densities, and at moderate dust temperatures. There are several wellstudied mechanisms that could cause this effect. Optical depth effects are most likely the major cause of this for the observed range of column densities (10 20-23 cm -2 ): CO for example, is thought to be optically thick in all but the lowest density gas within the Galaxy. An additional effect towards the high values of the column density would be CO depletion. In cold (T dust &lt; 20 K) and dense (n H 2 &gt; 10 4 cm -3 ) environments, CO molecules can 'freeze-out' on to the surfaces of dust grains forming icy mantles (e.g. <ref type="bibr">Fontani et al. 2012;</ref><ref type="bibr">Giannetti et al. 2014)</ref>, such that their observed gas-phase abundances are a factor of several smaller than expected in the case of no depletion (e.g. <ref type="bibr">Kramer et al. 1999;</ref><ref type="bibr">Hernandez et al. 2011)</ref>. We note, however, that we do not observe a strong dependence of the CO emission on the temperature at any given column density, and we would expect the freeze-out of CO to be strongly anticorrelated with increasing temperature, which we do not observe. This could be due to the large spatial scale of the beam used to study W49 (&#8764;3 pc).</p><p>Secondly, Fig. <ref type="figure">15</ref> shows that C 18 O has an emissivity that peaks at N H 2 &#8764; 10 21-22 cm -2 , which spans a temperature range of 15-25 K. As with the other CO isotopologues, above these densities, the C 18 O emission saturates, which is likely affected by the optical depth. Interestingly, here we find that for a given column density the emission efficiency ratio appears to peak towards higher dust temperatures (e.g. see between 10 21 and 2 &#215; 10 21 cm -2 ), which could tentatively point to the previously mentioned freeze-out of CO within the cold environments.</p><p>Thirdly, when analysing the emission efficiency ratios as a function of only the column density from Fig. <ref type="figure">14</ref>, we concluded that both HCN and N 2 H + peak at N H 2 &#8764; 3 &#215; 10 22 cm -2 ; an order of magnitude above the C 18 O peak. However, when we now compare with Fig. <ref type="figure">15</ref>, we find that around this peak in column density, the distribution of HCN and N 2 H + lines split into two discrete temperature peaks. We find that the N 2 H + peaks towards the lowest observed temperatures at the column density within the observed parameter space (&#8764;20 K), whereas the HCN peaks towards moderate-to-high temperatures of around &#8764;30-25 K. This then highlights the usefulness of the dust temperature as an additional variable to further investigate the emissivity profiles within the regions even at a low spatial resolution (&#8764;3 pc for these W49 observations). It is not surprising that both HCN and N 2 H + peak towards the densest gas as the abundance of N-bearing species is boosted within dense gas (e.g. <ref type="bibr">Caselli et al. 1999</ref>). However, it is somewhat surprising that they peak so clearly at different temperatures. None the less, we propose that this could be explained by the following. The abundance of N 2 H + is known to be particularly enhanced towards low temperatures, within regions where the CO molecules are frozen-out of the gas phase. This is due to CO being an effective destructor of the precursors in the formation pathway of N 2 H + ; removing molecules such as NH + and H + 3 . On the other hand, <ref type="bibr">Pety et al. (2017)</ref> showed that the molecules such as HCN (also HCO + , CN, and HNC) are sensitive to the far-UV radiation produced from star formation, where star-forming  <ref type="figure">2</ref>). We highlight by shading here column densities below the threshold limits for both sources, which are believed to suffer from foreground and background contamination (see Section 3.1). Values of the h Q below this threshold value should be taken with caution.</p><p>regions can be identified as having elevated dust temperatures. For example, the HCN (1-0) line is thought to be pumped by midinfrared photons and/or the effects of X-ray Dominated Regions (e.g. <ref type="bibr">Aalto et al. 2007)</ref>, which can both be present within actively star-forming regions (see Fig. <ref type="figure">1</ref>).</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.2">Focusing on the 'dense' gas within W49A</head><p>To better understand the line emission properties across the region we define a density model. To do so we focus on the central &#8764;40 pc (or 0.2 &#8226; ) of W49A,<ref type="foot">foot_5</ref> where we are the most confident that the majority of the molecular line emission and column density is associated to the material at or close to the distance of the W49A star-forming region (11.11 +0.79  -0.69 kpc; <ref type="bibr">Gwinn et al. 1992;</ref><ref type="bibr">Zhang et al. 2013)</ref>. As the true three-dimensional source structure of this region is not well understood, and there is almost certainly a significant level of sub-structure present on scales much smaller than the &#8764;3 pc beam size used throughout this work, we choose to define the most basic possible density model. We propose that the number density using the spatial resolution of the Herschel observations can also be obtained by limiting the three-dimensional structure to a single spatial element along the line of sight: or in other words, by assuming that the measured column density at each position originates from a 20 arcsec &#215; 20 arcsec &#215; 20 arcsec cube (or &#8764;1 pc 3 at &#8764;11 kpc). We determine an uncertainty on the model density by varying the cube size by (1) the linear size based on the uncertainty from the distance estimate, and (2) varying the depth of the box between 0.5 and 3 pc, which represent the average size of the sub-structures identified by <ref type="bibr">Eden et al. (2018)</ref>, and the size of the angular resolution of the molecular line observations (&#8764;60 arcsec), respectively.</p><p>The lower panel of Fig. <ref type="figure">16</ref> shows the distribution of the number density as a function of the observed column density. We find minimum, median, and maximum values for the model density across the W49A region of 1.9 &#215; 10 2 , 1.1 &#215; 10 3 , and 6.4 &#215; 10 4 cm -3 , respectively. The upper panel of Fig. <ref type="figure">16</ref> shows the cumulative distributions for a selection of the molecular lines as a function of column density when limited to the W49A region. We find that this plot recovers many of the trends seen in the cumulative distribution function of the entire mapped region (cf. Fig. <ref type="figure">9</ref>).</p><p>To then condense the cumulative distributions shown in Fig. <ref type="figure">16</ref> into a single value for each line, <ref type="bibr">Kauffmann et al. (2017)</ref> define a characteristic column density, N Q H 2 ,char , as the column density that contains half the total line intensity; W Q (N H 2 &lt; N H 2 ) = 50 per cent. Highlighted in Fig. <ref type="figure">16</ref> is the N Q H 2 ,char for CO, HCN, and N 2 H + , and the corresponding characteristic density:</p><p>). The characteristic column densities and densities for all lines are summarized within Table <ref type="table">4</ref>, which we can compare to the critical densities listed in Table <ref type="table">2</ref>. We find that the characteristic densities and critical densities for the low-density tracing CO lines are comparable (n Q char &#8764; 10 3 cm -3 ). However, we find that the higher density tracing molecular lines, such as HCN and N 2 H + , have characteristic densities only marginally higher than the CO lines (n Q char &#8764; 4 &#215; 10 3 cm -3 ), and hence significantly lower than their critical densities (n crit &#8764; 10 5-6 cm -3 ). A similar analysis was carried out for the Orion A molecular cloud by <ref type="bibr">Kauffmann et al. (2017, see their fig. 3)</ref>. Assuming a cylindrical density distribution, these authors determine a characteristic density for HCN of n Q char &#8776; 900 +1240 -550 cm -3 , which is comparable to within the uncertainty of the value calculated here (3400 &#177; 2800 cm -3 ).</p><p>Table <ref type="table">4</ref>. Properties determined from the cumulative distributions. Tabulated is the molecule, and its characteristic column density, N Q H 2 ,char , which is defined as the column density that contains half the total line intensity; W Q (N H 2 &lt; N H 2 ) = 50 per cent. The uncertainties show the range of N Q H 2 ,char after adding a synthetic noise of 1 &#963; W Q (see shaded region in Fig. <ref type="figure">9</ref>). We provide the results of the analysis across the whole mapped region (i.e. using the cumulative distributions shown in Section 4.1), and when limited to W49A (i.e. using the cumulative distributions shown in Section 5.2). Also shown for W49A is the characteristic density, which we define as n Q char = n(N Q H 2 ,char ) (see Section 5.2). This table has been ordered by increasing critical density (see Table <ref type="table">2</ref>). The full, machine-readable version of this table can be obtained from the supplementary online material.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Molecule</head><p>N Q</p><p>A potential explanation for the excess emission at characteristic densities below the critical density could be sub-thermal excitation (as previously mentioned for the extended map). Indeed, <ref type="bibr">Evans (1989)</ref> highlighted that the subthermal emission from gas below the critical density can be significant for molecular lines within the submm regime. Additionally, this could be partly a beam dilution effect. As the W49 region is known to contain significantly higher densities on smaller spatial scales that are not resolved by the observations presented in this work (e.g. <ref type="bibr">Vastel et al. 2001;</ref><ref type="bibr">Eden et al. 2018)</ref>. The line emission could then emit from these compact regions that are heavily beam diluted in the observations presented in this work. For lines such as HCN or N 2 H + , which have critical densities around three orders of magnitude higher than the observed characteristic density, we estimate that observations with a spatial resolution of around a factor of &#8730; 10 3 = 30 (i.e. 1 arcsec or 0.05 pc) would be required to investigate these beam dilution effects.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.3">Column density and mass conversions factors</head><p>There are two commonly adopted conversion factors within the literature that are used to estimate the mass of the material within a source from molecular line emission. Such conversions are useful for distant objects -such as extragalactic sources -where other methods of determining the gas mass of a system are not possible (e.g. dust extinction/continuum measurements). In this section, we investigate how the typically assumed canonical conversion factors adopted within the literature hold within the W49 region.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.3.1">Column density conversion; X-factor (X Q )</head><p>The first of these is typically referred to as the 'X-factor', and is used to convert the integrated intensity of a line to the molecular hydrogen column density. This can be defined as,</p><p>where X Q is the X-factor of a given molecular line transition in units of cm -2 (K km s -1 ) -1 . This conversion is typically used to estimate the total molecular gas mass from CO emission, which has an assumed canonical conversion value of X CO(1-0) = 2 &#215; 10 20 cm -2 (K km s -1 ) -1 <ref type="bibr">(Bolatto, Wolfire &amp; Leroy 2013)</ref>. Variations of the X CO(1-0) are, however, widely seen in both observations (see table <ref type="table">1</ref> of <ref type="bibr">Bolatto et al. 2013</ref>) and simulations (e.g. <ref type="bibr">Shetty et al. 2011a,b;</ref><ref type="bibr">Clark &amp; Glover 2015;</ref><ref type="bibr">Gong, Ostriker &amp; Kim 2018)</ref>. We calculate the X CO(1-0) factor using the mean molecular hydrogen column density across the mapped region, and the mean integrated intensity. Treating the whole mapped region as a single pixel, for this analysis we do not apply any uncertainty threshold clipping to the molecular line data set (i.e. X CO = N mean H 2 /W mean CO ; see Table <ref type="table">E2</ref> for the corresponding calculation for other lines). When doing so we find a mean X CO(1-0) = 0.3 &#215; 10 20 cm -2 (K km s -1 ) -1 across the mapped region (error on the mean shown as uncertainty), which is significantly below the value found within the literature <ref type="bibr">(Bolatto et al. 2013)</ref>. <ref type="bibr">Pety et al. (2017)</ref> found a similarly low value of X CO(1-0) towards the Orion B star-forming region. These authors attributed this difference to the UV illumination of the gas by massive stars, where only diffuse gas or the UV shielded dense gas have X CO(1-0) close to the typically assumed value. Manual inspection of X CO(1-0) on a pixel-by-pixel basis show elevated value towards the highest column density and temperature positions; towards the W49A region where embedded high-mass stars are known to reside. It is then difficult to attribute the X CO(1-0) discrepancy observed across W49 to UV illumination.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.3.2">Mass conversions; alpha-factor (&#945; Q )</head><p>The second conversion factor we investigate here is between the gas mass and the luminosity of the molecular line, &#945; Q . This can be defined as,</p><p>where &#945; Q is in units of M (K km s -1 pc 2 ) -1 . This conversion is typically used for molecular lines that are thought to trace only the dense gas, and hence to calculate the mass of dense gas within the region (e.g. using HCN; <ref type="bibr">Gao &amp; Solomon 2004a)</ref>. <ref type="bibr">Gao &amp; Solomon (2004b)</ref> outlined that the conversion for a virialized parcel of gas with a volume-averaged molecular hydrogen number density of n H2 &#8764; 3 &#215; 10 4 , and optically thick HCN (1-0) emission with a brightness temperature of T MB = 35 K, yields the value &#945; HCN(1-0) = 10M (K km s -1 pc 2 ) -1 . This calculation contains several fundamental assumptions for the HCN emission, which are not confirmed by the results presented in this work. For example, towards the centre of W49A we find peak main beam brightness temperatures for HCN of T MB &#8764; 4 K, and characteristic number densities of &#8764;10 3-4 cm -3 (Section 5.2), which would then give a significantly higher than standard value of &#945; HCN(1-0) &#8764; 15-50 M (K km s -1 pc 2 ) -1 [see Appendix E for the full derivation of &#945; HCN(1-0) ]. Indeed, higher values of the &#945; HCN(1-0) have been calculated for larger regions containing less efficiently emitting HCN (1-0). <ref type="bibr">Kauffmann et al. (2017)</ref>, for example, found &#945; HCN(1-0) = 20 M (K km s -1 pc 2 ) -1 for the Orion A molecular cloud (also see e.g. <ref type="bibr">Wu et al. 2010)</ref>. Most recently, <ref type="bibr">Evans et al. (2020)</ref> studied a sample of &#8764; 20 pc size molecular clumps within the Milky Way, and found &#945; HCN(1-0) = 20 M (K km s -1 pc 2 ) -1 within the gas above a visual extinction of A v &gt; 8 mag, and &#945; HCN(1-0) = 6 M (K km s -1 pc 2 ) -1 when integrating the HCN emission across the whole clumps. We calculate &#945; Q in both the total luminosity of the bulk molecular gas, and limiting the analysis to only the luminosity from the 'dense' molecular gas following the &#945; Q definitions of <ref type="bibr">Evans et al. (2020)</ref>. To do so, we firstly convert the integrated intensity maps to luminosity assuming a single distance of 11.11 kpc <ref type="bibr">(Zhang et al. 2013)</ref>. We then sum the luminosity across the whole mapped region without imposing an extinction threshold (L sum Q ). In addition, we determine the luminosity within a A v &gt; 8 mag contour (L sum Q,Av&gt;8 mag ), which is typically referred to as a limit for dense gas (e.g. see <ref type="bibr">Kauffmann et al. 2017)</ref>. We impose the same extinction threshold for the molecular gas (M sum Av&gt;8 mag ). The mass conversion is then calculated as the mass of A v &gt; 8 mag gas divided by the total luminosity of the given line:</p><p>This definition is most appropriate for determining the dense gas mass from unresolved (extra-galactic) observations that measure the total HCN luminosity including emission below and A v of 8 mag. Additionally, we calculate &#945; Q using both the mass and luminosity above the</p><p>Av&gt;8 mag /L sum Q,Av&gt;8 mag . This second method is appropriate for resolved (galactic) observations, where the HCN emission can be measured exclusively within the dense gas.</p><p>We find &#945; HCN(1-0) = 30 &#177; 1 M (K km s -1 pc 2 ) -1 , and <ref type="table">E1</ref> for additional conversion factors such as &#945; CO(1-0) ).<ref type="foot">foot_7</ref> These &#945; HCN(1-0) values are both significantly higher than the typically adopted value [&#8764;10-20 M (K km s -1 pc 2 ) -1 ], and more similar to the value determined analytically based on the observed source properties [15-50 M (K km s -1 pc 2 ) -1 ; Appendix E]. Fig. <ref type="figure">17</ref> shows a comparison of the measured &#945; HCN(1-0) to values available within the literature. The points have been coloured by extragalactic (blue; Garc&#237;a-Burillo et al. 2012), high-mass (red; <ref type="bibr">Evans et al. 2020;</ref><ref type="bibr">Nguyen-Luong et al. 2020)</ref>, and local star-forming regions (green; <ref type="bibr">Kauffmann et al. 2017;</ref><ref type="bibr">Shimajiri et al. 2017</ref>). We see that there is a large scatter in &#945; HCN(1-0) ranging &#8764;0.3-300 M (K km s -1 pc 2 ) -1 across all sources. Within this scatter, we find the high-mass and extragalactic sources appear to be systematically lower compared to local star-forming regions. We propose that the lower &#945; HCN(1-0) within higher mass regions could be due to the contribution of HCN emission from lower density gas present within subsequently large maps (i.e. towards more distance higher mass star-forming regions). Also highlighted on Fig. <ref type="figure">17</ref> is the map size used to determine each &#945; HCN(1-0) value, and indeed we see that higher mass starforming regions are studies across two to four orders of magnitude large areas compared to local star-forming regions. An alternative explanation, as suggested by <ref type="bibr">Shimajiri et al. (2017)</ref>, is that the higher levels of star formation and, hence, the radiation field could lower &#945; HCN(1-0) .</p><p>The analysis presented here highlights that significant caution should be exercised when adopting mass conversion factors from the literature. For example, if we take the mapped region as a typical bright 100 pc &#215; 100 pc pixel within an extragalactic observation, using the typically adopted dense gas conversion factor would cause an underestimation of the dense gas mass by a factor of up to &#8764;3 with respect to the extragalactic conversion factor &#945; HCN(1-0) determined in this work. This would cause an underestimation of the dense gas fraction (f dense = M dense /M tot ), and in turn, any property that contains f dense (e.g. the dense gas star formation efficiency: SFE dense = f dense M tot /SFR). This would then cause significant scatter in any comparison to the dense gas star formation efficiency relations (e.g. <ref type="bibr">Lada et al. 2010)</ref>.  <ref type="bibr">Burillo et al. 2012;</ref><ref type="bibr">Kauffmann et al. 2017;</ref><ref type="bibr">Shimajiri et al. 2017;</ref><ref type="bibr">Evans et al. 2020;</ref><ref type="bibr">Nguyen-Luong et al. 2020</ref>). The points have been coloured by the type of source studied by each work, which have been classified as extragalactic (blue), high-mass (red), and local star-forming regions (green). The horizontal dashed line shows the typically adopted value determined of &#945; HCN(1-0) = 10M (K km s -1 pc 2 ) -1 <ref type="bibr">(Gao &amp; Solomon 2004b)</ref>. Note that &#945; HCN(1-0) has been calculated for all sources using the mass of gas above an extinction threshold of A v 8 mag, divided by the total HCN luminosity across the region (&#945; Q (all) = M sum Av&gt;8 mag /L sum Q ). For each point, we label the size of the map used to determine both M sum Av&gt;8 mag and L sum Q . This figure highlights that the contribution of HCN emission from lower density gas present within large maps (i.e. towards more distance higher mass star forming regions) could be a potential cause for lower than typically assumed &#945; HCN(1-0) values. Additionally, the higher levels of star formation and, hence, radiation field have also been suggested to lower &#945; HCN(1-0) (see <ref type="bibr">Shimajiri et al. 2017)</ref>.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.4">Detection limits for extragalactic studies</head><p>In this section, we assume that the 100 pc &#215; 100 pc region of W49 studied here is representative of most massive star-forming regions within the Galaxy. We then pose a question: if this region were a single pixel in an extragalactic observation, would we be able to detect the molecular lines based on the brightnesses observed in this work? <ref type="bibr">Pety et al. (2017)</ref> performed a similar experiment for the Orion B star-forming region, yet this had a very limited spatial coverage of 5 pc &#215; 5 pc, which would correspond to an angular resolution at extragalactic distances that is currently difficult to obtain with even the largest mm-telescopes (e.g. ALMA; 1 pc at 10 Mpc is &#8764;0.02 arcsec). The observations presented here cover a much larger spatial area that corresponds to a projected angular size of &#8764;2 arcsec at 10 Mpc, which is becoming much more routinely observed over large areas of extragalactic systems (e.g. <ref type="bibr">Schinnerer et al. 2013;</ref><ref type="bibr">Leroy et al. 2016)</ref>.</p><p>We estimate that with 10 h of Atacama Large Millimeter Array observing time, we could reach a 3&#963; W Q &#8764; 0.3 K km s -1 line sensitivity within a &#8764;2 arcsec beam (100 pc at 10 Mpc) across a 200 arcsec &#215; 200 arcsec mosaic (or 10 kpc at 10 Mpc). <ref type="foot">10</ref>We compare this sensitivity to the mean integrated intensities across the mapped region listed in Table <ref type="table">3</ref>. We find that the CO (W mean Q &#8764; 80 K km s -1 ) and 13 CO (&#8764;10 K km s -1 ) molecular lines are significantly (&gt;100&#963; W Q ) above this value, and therefore would be easily detectable within the proposed ALMA observations. Additionally, lines such as HCN (1.4 K km s -1 ), C 18 O (1.0 K km s -1 ), and HCO + (0.9 K km s -1 ) are above this sensitivity estimate and should, therefore, also be moderately (10&#963; W Q ) detectable. To significantly detect e.g. N 2 H + (0.17 K km s -1 ), however, would require an increase in sensitivity of a factor of around two, which translates to a factor of around four more integration time (&#8764;46 hours).<ref type="foot">foot_9</ref> It is worth noting that these estimates have been determined using the W49 region, which is one of the brightest regions of molecular line emission within the Milky Way. Therefore, these estimate may not be representative of the detection expected across the whole disc for the proposed set of extragalactic observations.</p><p>In this section, we have discussed the potential to detect weak molecular lines (e.g. N 2 H + ) across the complete discs of nearby spiral galaxies thought to be comparable in their star formation properties to the Milky Way. However, it is worth highlighting that these lines have already been detected towards extragalactic sources using deep single-pointing or small mosaic observations (e.g. <ref type="bibr">Sage &amp; Ziurys 1995)</ref>. However, these detections are typically towards galactic centres, or starburst systems that are known to have extended and elevated e.g. HCN or N 2 H + emission relative to typical star-forming regions in disc galaxies (e.g. <ref type="bibr">Meier &amp; Turner 2005;</ref><ref type="bibr">Sakamoto et al. 2010)</ref>. For example, <ref type="bibr">Ginard et al. (2015)</ref> studied the central &#8764;kpc of the starburst galaxy M82 at scales of &#8764;50-100 pc, and found typical N 2 H + (1-0) integrated intensities in the range of &#8764;5-10 K km s -1 (see their fig. <ref type="figure">2</ref>). This is significantly higher than the mean N 2 H + (1-0) integrated intensity measured across the whole mapped region of W49 (see Table <ref type="table">3</ref>). The integrated intensity range of &#8764;5-10 K km s -1 is limited to the central &#8764;10 pc of W49A (see Fig. <ref type="figure">3</ref> and Table <ref type="table">B1</ref>), which may be indicative of the 'mini-starburst' nature of W49A (e.g. <ref type="bibr">Peng et al. 2010;</ref><ref type="bibr">Stock et al. 2014)</ref>.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="5.5">Caveats</head><p>At several points throughout the discussion, we mention the possible systematic uncertainties in the observations that may have affected our results. In this section, we summarize the most potentially significant of these caveats.</p><p>Limited spatial resolution: The observations presented throughout this work have been smoothed to a spatial resolution of &#8764;3 pc at the assumed source distance (11 kpc; <ref type="bibr">Zhang et al. 2013)</ref>. Higher resolution observations have shown that much sub-structure exists within W49A on scales much smaller than this beam size, which would then not be resolved with these observations <ref type="bibr">(Galv&#225;n-Madrid et al. 2013)</ref>. For example, using James Clerk Maxwell Telescope and higher resolution Herschel observations, <ref type="bibr">Eden et al. (2018)</ref> identified &#8764;50 compact sources within the central &#8764;60 pc of W49A that have effective radii ranging from 10 to 20 arcsec (&#8764;0.2 to 0.3 pc). These cores have masses of &gt;10 3 M , which when assuming a spherical geometry approximately have column densities of &gt;10 23 cm -2 . Column density values of 10 23 cm -2 are only observed towards the centre of W49A within the &#8764;60 arcsec smoothed Herschel maps (see Fig. <ref type="figure">1</ref>). The dense region identified by <ref type="bibr">Eden et al. (2018)</ref> would then be beam diluted within our smoothed maps, along with any molecular line emission originating from these regions. It is plausible that this would cause molecular line emission predominately originating from isolated compact, high (column) dense regions to appear to originate from low column densities when diluted. This is still, however, speculative and it is difficult to determine exactly how this would affect the results presented in this work.</p><p>We should note, however, that it is promising that the studies of the Orion star-forming regions find similar results to this work; e.g. emission efficiency ratio as a function of column density (section 8; <ref type="bibr">Kauffmann et al. 2017;</ref><ref type="bibr">Pety et al. 2017)</ref>. We previously mentioned that it is surprising that we even recover these trends, given that the spatial resolution of the Orion studies is &#8764;0.1 pc; an order of magnitude higher than this work (&#8764;3 pc). This might then highlight that limited spatial resolution is not such an issue, and it will be interesting to investigate if this holds for other distant sources within the LEGO survey.</p><p>Large spatial coverage: One of the fundamental improvements of this work, and the LEGO survey as a whole, over previous attempts to understand line emissivities of weak molecular line transitions is the large spatial coverage at high sensitivity. In the case of W49A, we have a total mapped area of &#8764;100 pc &#215; 100 pc at the source distance (11 kpc; <ref type="bibr">Zhang et al. 2013)</ref>. This has allowed us to analyse the molecular line emission across a very broad range of column densities. In this first work, we have considered the mapped region as a whole, and have not considered the individual environments contained within the map. These environments could, for example, differ considerably in their chemistry, and therefore have very different abundances of the various molecular species. Given the large distance to W49A, it is almost certain that some emission within the mapped area will not be physically associated and have very different chemical properties. For example, the W49B supernova remnant is within the mapped region (see Fig. <ref type="figure">1</ref>). However, it is not currently clear if W49B is in any way linked to W49A; with some surveys placing them up to several kilo parsecs away from each other <ref type="bibr">(Radhakrishnan et al. 1972;</ref><ref type="bibr">Chen et al. 2013;</ref><ref type="bibr">Zhu, Tian &amp; Zuo 2014)</ref>. One may assume that the physical and chemical properties of any molecular clouds interacting with this supernova could be different from those within actively star-forming region W49A.</p><p>Multiple sources along line of sight: This final caveat is somewhat linked to both the previous two, and concerns the fact that we do not separate the distinct sources along each line of sight within our map. We proposed that this may cause scatter in our results via two plausible scenarios: (a) if two (or more) dense clouds are present along with a single line of sight, (b) if one dense cloud is present with a large amount of foreground material. The first of these scenarios is most likely the case towards the centre of the mapped region for the majority of molecular lines. W49 is known to be a very kinematically complex region <ref type="bibr">(Galv&#225;n-Madrid et al. 2010)</ref>, and indeed manual inspection of the spectra show that multiple velocity components are present within this region. Away from the centre of W49A, the CO lines suffer the widespread overlap of velocity components at &#8764;10, &#8764;40, and &#8764;60 km s -1 , which are all extended across the mapped region. The remaining lines do not suffer from such large-scale spatial overlap; only peaking towards W49A at &#8764;10 km s -1 (albeit with several distinct components around this velocity), and then to the east of the mapped region at &#8764;60 km s -1 . An in-depth investigation of this velocity structure will be presented in future work. The second of the above-mentioned scenarios would cause an issue if there was a single dense cloud and a significant amount of diffuse foreground emission. This foreground material would contribute to the total column density along the line-of-sight, and to the more easily excited lines (e.g. CO). This material may not, however, emit strongly from e.g. HCN, which would cause an over (under) estimation in &#945; HCN (h HCN ) value.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head n="6">C O N C L U S I O N S</head><p>Molecular line emission has long been used to probe the properties of molecular clouds within the Milky Way, and now with the advent of the current generation of (sub)mm-telescopes, they are also quickly becoming major tools for the study of extragalactic systems. It is typically assumed that emission from carbon monoxide (CO) and its isotopologues allows us to probe the bulk gas properties in galaxies, whilst molecular species such as HCN and HCO + (and N 2 H + in future) potentially allow us to selectively characterize the densest gas. This is based on fundamental assumptions of the abundance and critical densities of these molecules across the density regimes within molecular clouds. It has, however, become clear such assumptions may not hold for all environments within individual, resolved star-forming regions (e.g. <ref type="bibr">Kauffmann et al. 2017;</ref><ref type="bibr">Pety et al. 2017;</ref><ref type="bibr">Evans et al. 2020)</ref>. For example, the so-called 'dense gas tracers' are assumed to emit at or above their critical densities (i.e. density when the collisional de-excitation rate is equal to the rate of spontaneous emission of photons for a molecular transition), however, additional mechanisms could play a major role in the total luminosity of the molecular line measured over a large region (e.g. <ref type="bibr">Evans 1999;</ref><ref type="bibr">Shirley 2015;</ref><ref type="bibr">Goldsmith &amp; Kauffmann 2017</ref>). Before we can exploit extragalactic line emission data to its full potential, we first need to develop a detailed understanding of these -and many other -emission lines in Milky Way molecular clouds. With this in mind, the 'Line Emission to assess Galaxy Observations' project (LEGO) aims at developing a first comprehensive picture of how the 3 mm band emission lines in Milky Way molecular clouds depend on factors like cloud density, star formation feedback, galactic environment, and metallicity.</p><p>In this paper, we present a study of emissivity of a selection of commonly observed molecular lines within the 3 mm band across a large mapped region containing the W49A massive starforming source (&#8764;100 pc &#215; 100 pc at 11 kpc). The main results are summarized below.</p><p>(i) We detect bright, extended emission from the CO and 13 CO (1-0) lines, which have significant emission -i.e. pixels with integrated intensities higher than three times their associated uncertainty -across the entire mapped region (Section 4.1). The C 18 O (1-0) is the next most extended line, which has significant emission over half the mapped region (spatial coverage of &#8764;35 per cent). We find that the higher density tracing lines such as HCN, HCO + (1-0) cover just under a third of the mapped region (spatial coverage of &#8764;30 per cent), and, therefore, are fairly extended yet much less than CO and 13 CO (1-0). Lastly, the lines such as N 2 H + (1-0) are the most spatially compact (spatial coverage of &lt; 10 per cent). These results show the critical density of the observed molecular lines cannot be used as trivial predictors for how extended the emission is across a star-forming region (Fig. <ref type="figure">6</ref>). Such a result is not surprising and could be easily explained by the lower abundance and, therefore, optical depth, of rarer molecules. In other words, it would not be fair to suggest that molecular line X should be more extended or brighter than molecular line Y across a given parcel of molecular gas solely because of the lower critical density of molecular line Y. Rather, the relative abundances of the molecules X and Y should be taken into account. As if the molecular Y is significantly more abundant than molecule X, molecular line Y may be brighter and more extended than molecular line X regardless of its lower critical density.</p><p>(ii) It is desirable to distil the emission characteristics of each molecular line into a single number. To do so, following the analysis of <ref type="bibr">Kauffmann et al. (2017)</ref>, we define the characteristic column density at which half the line intensity is observed [N Q H 2 ,char = N H 2 (W Q = 50 per cent); Section 4.1]. We find that the dense gas tracers (e.g. HCN and N 2 H + ) have higher N Q H 2 ,char relative to the CO lines. We also outline a simple number density model for the central region of W49 (Section 5.2), which we then use to define the characteristic density for each line</p><p>. We find that the CO lines have comparable characteristic densities and critical densities (n Q char &#8764; 10 3 cm -3 ). However, the higher density tracing molecular lines, such as HCN and N 2 H + , have characteristic densities only marginally higher than the CO lines (n Q char &#8764; 4 &#215; 10 3 cm -3 ), and hence significantly lower than their critical densities (n crit &#8764; 10 5-6 cm -3 ). A potential explanation for the excess emission at characteristic densities below the critical density could be sub-thermal excitation, which can be significant for molecular lines within the sub-mm regime (e.g. <ref type="bibr">Evans 1989</ref>). This confirms that the critical density of a line cannot be used as a probe of the physical volume density in an absolute sense. That said, the use of emission from dense gas tracers, such as HCN and N 2 H + , relative to CO emission may still be a useful tool to probe the proportion of relatively denser gas. Future works from the LEGO survey will aim at expanding the dynamic range in n Q char between CO and the dense gas tracers.</p><p>(iii) The ratio of the integrated intensity over the column density can be thought as a proxy for the line emissivity, or the efficiency of an emitting transition per H 2 molecule (also previously called the emission efficiency). One may assume that the emissivity is a function of number density, and, therefore, a function of column density, which would be the case if the molecular abundance or excitation is directly linked to the density <ref type="bibr">(Kauffmann et al. 2017)</ref>. We find, however, significant variations in the emissivity of the various molecular line transitions (Section 3.3). The CO and 13 CO transitions show an overall decrease with increasing column density (see Fig. <ref type="figure">8</ref>). The C 18 O transition shows a peak at &#8764;10 21 cm -2 , and decreases towards higher column density values. We find that the remaining lines of HCN, HCO + , HNC, CN, CS, CH 3 OH, and N 2 H + have similar peaked profiles, albeit towards higher values of the column density (few 10 22 cm -2 ). These trends are broadly in agreement with the results found towards the Orion star-forming regions (e.g. <ref type="bibr">Kauffmann et al. 2017;</ref><ref type="bibr">Pety et al. 2017)</ref>, which is somewhat surprising given the observations presented here have over an order of magnitude larger spatial resolution (&#8764;2-3 pc at the source distance of 11 kpc).</p><p>(iv) Direct comparison to the Orion A results <ref type="bibr">(Kauffmann et al. 2017)</ref>, however, highlight several relative differences between the emission efficiency profiles between the two regions (see Fig. <ref type="figure">14</ref>). For example, within Orion the HCN emission efficiency peaks at a lower column density than N 2 H + , whereas in W49 we find these are comparable, with N 2 H + even potentially peaking at lower column density values. To further investigate this, we assess how the emission efficiency at a given column density varies as a function of dust temperature (Section 4.2). We find that around the single column density emission efficiency peak, the distribution for HCN and N 2 H + lines split into two discrete emission efficiency temperature peaks. We find that the N 2 H + peaks towards the lowest observed temperatures at the highest column density within the observed parameter space (&#8764;20 K), whereas the HCN peaks towards moderateto-high temperatures of around &#8764;30-25 K. We believe this could be caused by the chemical formation pathways of these molecules that favour different temperature regimes, or the enhancement in the emission due to transition specific excitation mechanisms [e.g. the mid-IR pumping of HCN (1-0)].</p><p>(v) We also investigate several commonly used conversion factors between molecular line emission and molecular gas mass and column density across the W49 region (Section 5.3). First, we focus on the conversion factor between the integrated intensity of CO and column density (N H 2 = X Q W Q ; Section 5.3). We calculate an average value across the mapped region of X CO(1-0) = 0.297 &#177; 0.006 &#215; 10 20 cm -2 (K km s -1 ) -1 , which is significantly lower than the canonical value within the literature [&#8764;2 &#215; 10 20 cm -2 (K km s -1 ) -1 ; <ref type="bibr">Bolatto et al. 2013]</ref>. Secondly, we investigated the conversion between the total luminosity of HCN and gas mass above a visual extinction of A v &gt; 8 mag within the region (i.e. &#945; Q = M sum Av&gt;8 mag /L sum Q ). We find</p><p>which is significantly higher than the typically adopted value within the literature [10 M (K km s -1 pc 2 ) -1 ; <ref type="bibr">Gao &amp; Solomon 2004a,b]</ref>.</p><p>If we assume that the 100 pc &#215; 100 pc region of W49 studied here is a typical massive star-forming complex within the Milky Way, or within an extragalactic system, these results have serious implications for our understanding of star formation. For example, if we take our mapped region as a single 100 pc &#215; 100 pc pixel in an extragalactic survey, and use the typically assumed conversion factors as opposed to the conversion factors presented in this work. We would overestimate the total molecular column density and molecular gas mass by a factor of seven, and underestimate the dense gas mass by a factor of three. This would then cause over an order of magnitude underestimation of a dense gas mass fraction (f dense = M dense /M tot ), and in turn, cause a large systematic offset in any property that contains f dense (e.g. the dense gas mass star formation efficiency: SFE dense = f dense M tot /SFR). It is worth noting that here we take the &#945; HCN(1-0) measured over the mapped region as an example single pixel in an extragalactic observation, and, in reality, the statistics achieved with large maps of extragalactic systems may cause deviations in &#945; HCN(1-0) to be averaged-out over various mass and evolutionary stage star-forming regions. None the less, this example should highlight the need for further resolved measurements of mass conversion factors to precisely measure such properties as the dense gas star formation efficiency, which are needed to accurately constrain current star formation theories.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>AC K N OW L E D G E M E N T S</head><p>We would like to thank the referee for their constructive feedback that helped improve the paper </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>online.zip</head><p>Please note: Oxford University Press is not responsible for the content or functionality of any supporting materials supplied by the authors. Any queries (other than missing material) should be directed to the corresponding author for the article.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A P P E N D I X A : L E G O M O L E C U L A R L I N E DATA BA S E I N F O R M AT I O N</head><p>Tables <ref type="table">2</ref> and<ref type="table">A1</ref> record the information on molecular emission lines the LEGO survey uses to observe and characterize emission lines. Line frequencies are from <ref type="bibr">Lovas (2004)</ref>, while energy levels, Einstein coefficients and downward collisional rate coefficients are taken from the Leiden Atomic and Molecular Database <ref type="bibr">(Sch&#246;ier et al. 2005</ref>; LAMDA, accessed in November 2019). Section 2.1.2 provides further information on how Tables <ref type="table">2</ref> and<ref type="table">A1</ref> were constructed.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A P P E N D I X B : I N T E G R AT E D I N T E N S I T Y M A P S TAT I S T I C S</head><p>In this section, we present statistics determined from the molecular line integrated intensity maps (see Section 2.1.3). Shown in Table <ref type="table">B1</ref> is the minimum value of all pixels across each map (i.e. without taking a significance threshold). We see that CO and 13 CO have positive minimum values, which shows that they are detected across all positions within the map (also see Fig. <ref type="figure">3</ref>). The other molecular lines, however, are negative, showing that they are detected above a threshold intensity. This is given as the minimum value above W Q &gt; 3&#963; W Q also given in Table <ref type="table">B1</ref>. In addition, we also give the minimum, 5, 16, 50, 84, and 95 percentile ranges when taking only the significant positions. Note that Table <ref type="table">B1</ref> has been ordered by decreasing A cov values (see Table <ref type="table">3</ref>). We can see that there is little correlation to the line brightness and its coverage across the region (e.g. see W max Q ). Some lines are, therefore, seen as extended as relative faint [e.g. CCH (1-0, 1/2-1/2)], whilst other are seen as compact and bright [e.g. SiO (2-1); see Fig. <ref type="figure">7</ref>].</p><p>Table <ref type="table">A1</ref>. Details on the emission lines covered by the LEGO survey, and processed by our pipeline. Section 2.1.2 explains how lines examined in this study were selected. Each of the selected lines was given a LEGO reference code in the reduction pipeline, which are presented in the first column of the table. The rest frequency recorded in this table, &#957; rest , refers to the specific full transition <ref type="bibr">(Lovas 2004</ref>). These transitions are sometimes part of larger groups, as explained in Section 2.1.2 (see the final column of this table). In those cases, &#957; min and &#957; max state the minimum and maximum frequencies of the lines considered to form a group. The full, machine-readable version of this table can be obtained from the supplementary online material.  </p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A P P E N D I X C : R A N D O M U N C E RTA I N T I E S O N T H E H 2 C O L U M N D E N S I T Y A N D D U S T T E M P E R AT U R E M A P S</head><p>We follow the weighted least squared fitting procedure outlined in <ref type="bibr">Guzm&#225;n et al. (2015)</ref> to calculate the molecular hydrogen column densities and dust temperatures from the Herschel Hi-Gal observations (see Section 2). This involves fitting the Herschel filter fluxes to the modified blackbody model, and adjusting the free parameters (surface density and dust temperature) in order to minimize the weighted least squared difference:</p><p>where i runs through each filter, &#963; 2 i = (0.1 * data i ) 2 + &#963; 2 i,N and &#963; i,N is given by the fifth column of table 1 in <ref type="bibr">Guzm&#225;n et al. (2015)</ref>.</p><p>To get the random uncertainty associated with the fit at each position, we follow the method outlined by Lampton, Margon &amp; Bowyer <ref type="bibr">(1976)</ref>. We test how 'sensitive' is the minimization function to variations in the input parameters. In practice, this means finding the &#967; 2 = &#967; 2 min + 2.3 contour (2.3 is adequate to get 1&#963; confidence intervals with two free parameters). This contour (usually close to an ellipse) defines an uncertainty region, whose projections on to the parameter axes define the 1&#963; confidence intervals (see fig. <ref type="figure">3</ref> in <ref type="bibr">Guzm&#225;n et al. 2015)</ref>. Fig. <ref type="figure">C1</ref> shows the size of these confidence intervals towards the positive and negative side of the best-fitting parameter. The parameters shown here are the dust temperature, and the logarithm of the H 2 column density.</p><p>We investigate the fractional uncertainty across the mapped region by taking the ratios of the positive and negative uncertainty at each position over the corresponding measured value at the same position. The distribution of the H 2 column density and dust temperature fractional uncertainties are plotted in Fig. <ref type="figure">C2</ref>. We find that the mean fractional uncertainties produced by random noise on the column density and dust temperature maps are In this section, we present a brief inspection of the velocity structure observed within the molecular line data cubes. To do so, we integrate the line intensity within discrete velocity windows across the velocity range where the significant emission is seen in the average spectra (Fig. <ref type="figure">2</ref>). Fig. <ref type="figure">D1</ref> displays the result when using velocity windows of -20 to 80 km s -1 in steps of 20 km s -1 for the CO (1-0), C 18 O (1-0), and HCN (1-0) maps. We note that to produce these maps we use cubes that have all spectral channels below 5 &#963; rms masked (see Table <ref type="table">2</ref>), rather than using the masked cubes produced by the CO masking routine (see Section 2.1.3). These plots highlight the three broad velocity components that are present across the mapped region, which peak at velocities of &#8764;10, &#8764;40, &#8764;60 km s -1 . The brightest of these is the &#8764;10 km s -1 component, which is seen in all the displayed lines and appears to be spatially correspondent with the W49A star-forming region (see Fig. <ref type="figure">1</ref>). Table <ref type="table">E1</ref>. Calculated conversion factors for the various molecular line transitions (see Section 5.3). We highlight the two most commonly adopted conversion factors: X CO(1-0) and &#945; HCN(1-0) . The values shown in the upper half of the table have been calculated using the column density and mass of A v &gt; 8 mag gas divided by the mean integrated intensity and total luminosity of the given line across the region (e.g. &#945; Q = M sum Av&gt;8 mag /L sum Q ), whilst those within the lower half have been calculated using both the column density and mass, and integrated intensity and luminosity above the A v &gt; 8 mag threshold (&#945; Q (A v &gt; 8 mag) = M sum Av&gt;8 mag /L sum Q,Av&gt;8 mag ). The full, machine-readable version of this table can be obtained from the supplementary online material.</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>A P P E</head></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Line log(W mean</head><p>( Kk ms -1 pc 2 ) [ c m -2 (K km s -1 ) -1 ] [ M (K km s -1 pc 2 ) -1 ] (&#215; 10 20 )  <ref type="table">E1</ref>). The values shown have been calculated without imposing any extinction threshold on the gas column density and gas mass, and the integrated intensity and luminosity of the given line (such that e.g. &#945; Q = M sum /L sum Q ). The full, machine-readable version of this table can be obtained from the supplementary online material. and substituting this into equation (E2),</p></div>
<div xmlns="http://www.tei-c.org/ns/1.0"><head>Line log(W mean</head><p>The luminosity of a molecular line transition, L, can be approximated as the product of the integrated intensity (W = T MB &#8730; 2&#963; ), and the area over which the intensity has been measured (A = &#960; R 2 ),</p><p>Substituting equations (E4) and (E5) in the above gives, L &#8733; T MB &#961; -2.5 . (E8)</p><p>Finally, taking the ratio of the virial mass in equation (E6), and the luminosity in equation (E8) yields, &#945; = ML -1 &#8733; &#961; -0.5 (T MB &#961; -2.5 ) -1 &#8733; &#961; 0.5 T MB , (E9) which in terms of molecular hydrogen number density, n H 2 , can be approximated as,</p><p>Gao <ref type="bibr">&amp; Solomon (2004a,b)</ref> approximate that a virialized core with a volume averaged molecular hydrogen number density of n H 2 &#8764; 3 &#215; 10 4 cm -3 , and optically think HCN (1-0) emission with a brightness temperature of T MB = 35 K, gives &#945; HCN(1-0) = 10 M (K km s -1 pc 2 ) -1 . Alternatively, in this work we measure a peak main beam brightness temperatures for HCN of T MB &#8764; 4 K, and characteristic number densities of &#8764;10 3-4 cm -3 (Section 5.2), which would then give a significantly higher than standard values of &#945; HCN(1-0) &#8764; 15 to -50 M (K km s -1 pc 2 ) -1 [see Appendix E for the full derivation of &#945; HCN(1-0) ]. The column density conversion X Q = W Q /N H 2 follows a very similar calculation, and can be found in full in e.g. <ref type="bibr">Dickman, Snell &amp; Schloerb (1986)</ref>.</p></div><note xmlns="http://www.tei-c.org/ns/1.0" place="foot" xml:id="foot_0"><p>MNRAS 497,1972-2001 (2020)    Downloaded from https://academic.oup.com/mnras/article/497/2/1972/5880778 by MIT Libraries user on 27 August 2020</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="2" xml:id="foot_1"><p>http://www.iram.es/IRAMES/mainWiki/Iram30mEfficiencies</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="3" xml:id="foot_2"><p>http://www.iram.fr/IRAMFR/GILDAS</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="4" xml:id="foot_3"><p>https://www.iram-institute.org/EN/content-page-247-7-55-177-247-0.ht ml MNRAS 497, 1972-2001 (2020) Downloaded from https://academic.oup.com/mnras/article/497/2/1972/5880778 by MIT Libraries user on 27 August 2020</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="6" xml:id="foot_4"><p>It is worth considering that there are several (typically weaker) CH 3 OH transitions within the frequency range that we have not considered within this work. These include the 2(0)-1(0) and 2(1)-1(1) transitions of CH 3 OH-E at rest frequencies of 96.744 545 GHz and 96.755 501 GHz, respectively (e.g.<ref type="bibr">Menten et al. 1988;</ref><ref type="bibr">Leurini et al. 2004</ref>). This is noted within TableA1within the Appendix of this work. MNRAS 497, 1972-2001 (2020) Downloaded from https://academic.oup.com/mnras/article/497/2/1972/5880778 by MIT Libraries user on 27 August 2020</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="8" xml:id="foot_5"><p>We choose to define the central position of the W49A region as RA, Dec. (J2000) = 19 h 10 m 14 s ,</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="9" xml:id="foot_6"><p>&#8226; 06 &#8226; 17 , and take only the pixels within a box of 0.2 &#8226; around this position for the analysis presented in this section.</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="9" xml:id="foot_7"><p>Table E2 provides &#945; Q values that have been determined without imposing extinction thresholds on either the mass or luminosity measurements. MNRAS 497, 1972-2001 (2020) Downloaded from https://academic.oup.com/mnras/article/497/2/1972/5880778 by MIT Libraries user on 27 August 2020</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="10" xml:id="foot_8"><p>This time estimate has been performed using the ALMA Cycle 8 observing tool, and accounts for only the (C43-3 configuration) 12 m array observations. For these estimates, we choose the rest frequency of the HCN(1-0) transition, a representative bandwidth of 282 kHz (or 0.96 km s -1 ), and Nyquist sampling of 39 12 m array pointings that create a 200 arcsec &#215; 200 arcsec mosaic. The quoted time includes the on-source time to reach 3&#963; W Q &#8764; 0.3 K km s -1 of around 7 h, and calibration and overhead time of around 4.12 h.</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="11" xml:id="foot_9"><p>The on-source only</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" n="12" xml:id="foot_10"><p>m array observing time required to reach 3&#963; W Q &#8764; 0.15 K km s -1 would be 27.6 h. MNRAS 497, 1972-2001 (2020) Downloaded from https://academic.oup.com/mnras/article/497/2/1972/5880778 by MIT Libraries user on 27 August 2020</p></note>
			<note xmlns="http://www.tei-c.org/ns/1.0" place="foot" xml:id="foot_11"><p>This paper has been typeset from a T E X/L A T E X file prepared by the author.MNRAS 497,1972-2001 (2020)    Downloaded from https://academic.oup.com/mnras/article/497/2/1972/5880778 by MIT Libraries user on 27 August 2020</p></note>
		</body>
		</text>
</TEI>
