arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:1809.03378v1 [cs.IT] 10 Sep 2018

Machine Learning Based Hybrid Precoding for MmWave MIMO-OFDM with Dynamic Subarray

Yiwei Sun1, Zhen Gao2, Hua Wang1, and Di Wu3 Affiliation: 1School of Information and Electronics, Beijing Institute of Technology, Beijing, China Affiliation: 2Advanced Research Institute of Multidisciplinary Science, Beijing Institute of Technology, Beijing, China Affiliation: 3China Academy of Information and Communications Technology, Beijing, China
{sunyiwei, gaozhen16, wanghua}@bit.edu.cn
Abstract

Hybrid precoding design can be challenging for broadband millimeter-wave (mmWave) massive MIMO due to the frequency-flat analog precoder in radio frequency (RF). Prior broadband hybrid precoding work usually focuses on fully-connected array (FCA), while seldom considers the energy-efficient partially-connected subarray (PCS) including the fixed subarray (FS) and dynamic subarray (DS). Against this background, this paper proposes a machine learning based broadband hybrid precoding for mmWave massive MIMO with DS. Specifically, we first propose an optimal hybrid precoder based on principal component analysis (PCA) for the FS, whereby the frequency-flat RF precoder for each subarray is extracted from the principle component of the optimal frequency-selective precoders for fully-digital MIMO. Moreover, we extend the PCA-based hybrid precoding to DS, where a shared agglomerative hierarchical clustering (AHC) algorithm developed from machine learning is proposed to group the DS for improved spectral efficiency (SE). Finally, we investigate the energy efficiency (EE) of the proposed scheme for both passive and active antennas. Simulations have confirmed that the proposed scheme outperforms conventional schemes in both SE and EE.

Index Terms: 
Hybrid precoding, MIMO-OFDM, millimeter wave, machine learning, dynamic subarray, energy efficiency.

I Introduction

In the fifth generation mobile communications, the application of millimeter-wave (mmWave) is vital by the virtue of providing high data rate and large bandwidth [1, 2, 3]. However, mmWave channel suffers from a severe path loss, and the traditional fully-digital precoding with massive antennas to mitigate this issue is extremely power consuming [4]. Therefore, hybrid precoding has been proposed to achieve the large array gains with the reduced hardware cost and power consumption [5, 6, 7, 8]. By far, existing broadband hybrid precoding schemes usually focus on fully-connected array (FCA), while seldom consider partially-connected subarray (PCS) like fixed subarray (FS) and dynamic subarray (DS). Therefore, hybrid precoding with PCS in broadband channel is an interesting topic to explore.

Most prior work is based on narrowband mmWave channels [11, 10, 9]. Specifically, a compressive sensing-based hybrid precoding has been proposed in [11], where the channel sparsity is ingeniously exploited to design hybrid precoding with the aid of orthogonal matching pursuit (OMP) algorithm. Moreover, a constant envelope hybrid precoding scheme is proposed, where two cost-efficient sub-connected hybrid architectures are considered to optimize the hybrid precoding under per-antenna constant envelope constraints [10]. To improve bit-error-rate, an over-sampling codebook-based hybrid minimum sum-mean-square-error precoding is designed [9]. On the other hand, mmWave channels appear to have the frequency-selective fading, where OFDM is usually adopted to combat the time dispersion channels [12, 13, 14]. Specifically, an insightful broadband hybrid precoder based on limited-feedback codebook has been proposed for FCA [12]. By exploiting the channel correlation information among different subcarriers, a broadband hybrid precoding is proposed for FS and DS [13]. Finally, [14] has theoretically shown the optimality of frequency flat precoding by proving that dominant subspaces of the frequency domain channel matrices of different subcarriers are equivalent. However, this conclusion is based on the purely sparse channels with discrete angles of arrival (AoA) and angles of departure (AoD), while the explicit precoder solution is not provided.

In this paper, we propose a machine learning based hybrid precoding scheme for mmWave MIMO-OFDM systems with DS, where a shared agglomerative hierarchical clustering (shared-AHC) algorithm is proposed for DS grouping to improve SE performance. First, we propose a PCA-based analog precoder scheme for FS by abstracting the low-dimensional signal space of frequency-flat precoder for given subarray from the high-dimensional signal space of optimal frequency-selective fully-digital precoders using PCA. Besides, the optimality of the proposed PCA-based hybrid precoder design is theoretically proven and verified by simulations. Second, we propose the shared-AHC algorithm inspired by cluster analysis in the field of machine learning for antenna grouping. By implementing shared-AHC algorithm, the SE performance of PCS can be further enhanced for effective antenna grouping adapting to the spatial features of the frequency-selective channels. Finally, we consider the practical passive/active antennas for EE performance analysis. Simulation results confirm the better spectral efficiency (SE) and energy efficiency (EE) performance achieved by the proposed scheme than existing schemes. Meanwhile, DS has the overwhelming advantage for both active and passive antennas.

Notations: Following notations are used throughout this paper. 𝐀\mathbf{A} is a matrix, 𝐚\mathbf{a} is a vector, aa is a scalar, and 𝒜\mathcal{A} is a set. Conjugate transpose and transpose of 𝐀\mathbf{A} are 𝐀H\mathbf{A}^{H} and 𝐀T\mathbf{A}^{T}, respectively. The (i,j)(i,j)th entry of 𝐀\mathbf{A} is [𝐀]i,j[\mathbf{A}]_{i,j}, and [𝐀]i,:[\mathbf{A}]_{i,:} ([𝐀]:,j[\mathbf{A}]_{:,j}) denotes the iith row (jjth column) of 𝐀\mathbf{A}. Frobenius norm is denoted by ||||F||\cdot||_{F}. |𝐀||\mathbf{A}|, |𝒜||\mathcal{A}|, |𝐚||\mathbf{a}|, and |a||a| are the determinant of a square matrix 𝐀\mathbf{A}, cardinality of a set 𝒜\mathcal{A}, 2\ell_{2}-norm of a vector 𝐚\mathbf{a}, and modulus of a number aa, respectively. The iith largest singular value of a matrix 𝐀\mathbf{A} is defined as λi(𝐀)\lambda_{i}(\mathbf{A}). Additionally, blkdiag(𝐚1,,𝐚K)\text{blkdiag}(\mathbf{a}_{1},\cdots,\mathbf{a}_{K}) is a block diagonal matrix with 𝐚i\mathbf{a}_{i} (1iK1\leq i\leq K) on its diagonal blocks.

II System Model

We consider an mmWave massive 3-dimensional (3D) MIMO system, where both the BS and user employ the uniform planar array (UPA), and OFDM is adopted to combat the frequency-selective fading channels. The BS is equipped with Nt=Ntv×NthN_{t}=N_{t}^{v}\times N_{t}^{h} antennas and NtRFNtN_{t}^{\rm RF}\ll N_{t} chains, where NtvN_{t}^{v} and NthN_{t}^{h} are the numbers of vertical and horizontal transmit antennas, respectively. The user is equipped with Nr=Nrv×NrhN_{r}=N_{r}^{v}\times N_{r}^{h} antennas, where NrvN_{r}^{v} and NrhN_{r}^{h} are the numbers of vertical and horizontal receive antennas, respectively. We consider the downlink transmission, and the received symbols at the user can be written as [11]

𝐫[k]=𝐖H[k](𝐇[k]𝐅RF𝐅BB[k]𝐱[k]+𝐧[k]),\mathbf{r}[k]=\mathbf{W}^{H}[k](\mathbf{H}[k]\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]\mathbf{x}[k]+\mathbf{n}[k]), (1)

where 1kK1\leq k\leq K with KK being the number of subcarriers, 𝐅BB[k]NtRF×Ns\mathbf{F}_{\rm BB}[k]\in\mathbb{C}^{N_{t}^{\rm RF}\times N_{s}}, 𝐅RFNt×NtRF\mathbf{F}_{\rm RF}\in\mathbb{C}^{N_{t}\times N_{t}^{\rm RF}}, 𝐖[k]Nr×Ns\mathbf{W}[k]\in\mathbb{C}^{N_{r}\times N_{s}}, 𝐇[k]Nr×Nt\mathbf{H}[k]\in\mathbb{C}^{N_{r}\times N_{t}}, 𝐱[k]Ns×1\mathbf{x}[k]\in\mathbb{C}^{N_{s}\times 1}, and 𝐧[k]Nr×1\mathbf{n}[k]\in\mathbb{C}^{N_{r}\times 1} are the digital precoder, analog precoder, fully-digital combiner, channel matrix, transmitted signal, and noise associated with the kkth subcarrier, respectively, and NsN_{s} is the number of data streams for each subcarrier. Noise 𝐧[k]\mathbf{n}[k] satisfies 𝐧[k]𝒞𝒩(0,σn2)\mathbf{n}[k]\sim\mathcal{CN}(0,\sigma_{n}^{2}), and transmitted signal 𝐱[k]\mathbf{x}[k] satisfies 𝔼[𝐱[k]𝐱H[k]]=PKNs\mathbb{E}[\mathbf{x}[k]\mathbf{x}^{H}\![k]]\!\!=\!\!\frac{P}{KN_{s}}, where PP is average total transmit power.

The frequency-domain channel 𝐇[k]\mathbf{H}[k] can be expressed as 𝐇[k]=d=0D1𝐇d[d]ej2πkKd\mathbf{H}[k]=\sum_{d=0}^{D-1}\mathbf{H}_{d}[d]e^{-j\frac{2\pi k}{K}d} [12], where DD is the maximum delay spread of the discretized channels, and 𝐇d[d]Nr×Nt\mathbf{H}_{d}[d]\in\mathbb{C}^{N_{r}\times N_{t}} is the delay-dd channel matrix. We consider the clustered channel model [11], where the channel is composed by NclN_{\rm cl} clusters of multipaths with NrayN_{\rm ray} rays in each cluster. Thus the delay-dd channel matrix can be written as

𝐇d[d]=i=1Ncll=1Nraypi,ld[d]𝐚r(ϕi,lr,θi,lr)𝐚tH(ϕi,lt,θi,lt),\mathbf{H}_{d}[d]=\sum\nolimits_{i=1}^{N_{\rm cl}}\sum\nolimits_{l=1}^{N_{\rm ray}}p_{i,l}^{d}[d]\mathbf{a}_{r}(\phi^{r}_{i,l},\theta^{r}_{i,l})\mathbf{a}_{t}^{H}(\phi^{t}_{i,l},\theta^{t}_{i,l}), (2)

where pi,ld[d]=NtNr/(NclNray)αi,lp(dTsτi,l)p_{i,l}^{d}[d]=\sqrt{N_{t}N_{r}/(N_{\rm cl}N_{\rm ray})}\alpha_{i,l}p(dT_{s}-\tau_{i,l}) is the delay-domain channel coefficient, τi,l\tau_{i,l}, αi,l\alpha_{i,l}, and p(τ)p(\tau) are the delay, the complex path gain, and the pulse shaping filter for TsT_{s}-spaced signaling, respectively. Thus the relationship between the frequency-domain channel coefficiency and the delay-domain channel coefficiency is pi,l[k]=d=0D1pi,l[d]exp(j2πkd/K)p_{i,l}[k]=\sum_{d=0}^{D-1}p_{i,l}[d]\exp(-j2\pi kd/K). In (2), 𝐚t(ϕi,lt,θi,lt)\mathbf{a}_{t}(\phi^{t}_{i,l},\theta^{t}_{i,l}) and 𝐚r(ϕi,lr,θi,lr)\mathbf{a}_{r}(\phi^{r}_{i,l},\theta^{r}_{i,l}) are the steering vectors of the llth path in the iith cluster at the transmitter and receiver, respectively. In the steering vectors, ϕi,lt\phi^{t}_{i,l} and θi,lt\theta^{t}_{i,l} are the azimuth and elevation angles of the llth ray in the iith cluster for AoDs, and ϕi,lr\phi^{r}_{i,l} and θi,lr\theta^{r}_{i,l} are the azimuth and elevation angles of the llth ray in the iith cluster for AoAs. Therefore, the transmit steering vectors for the UPA at the BS can be expressed as 𝐚t(ϕi,lt,θi,lt)=[1ej2π(mdhλsin(θi,lt)cos(ϕi,lt)+ndvλsin(ϕi,lt))ej2π((Nth1)dhλsin(θi,lt)cos(ϕi,lt)+(Ntv1)dvλsin(ϕi,lt))]T/Nt\mathbf{a}_{t}(\phi^{t}_{i,l},\theta^{t}_{i,l})=[1\ \ \ \cdots\ \ \ e^{-j2\pi(m\frac{d_{h}}{\lambda}\sin(\theta^{t}_{i,l})\cos(\phi^{t}_{i,l})+n\frac{d_{v}}{\lambda}\sin(\phi^{t}_{i,l}))}\ \ \ \cdots\\ e^{-j2\pi((N_{t}^{h}-1)\frac{d_{h}}{\lambda}\sin(\theta^{t}_{i,l})\cos(\phi^{t}_{i,l})+(N_{t}^{v}-1)\frac{d_{v}}{\lambda}\sin(\phi^{t}_{i,l}))}]^{T}/\sqrt{N_{t}} [11], where λ\lambda is the carrier wavelength, and dvd_{v} and dhd_{h} are the distances between adjacent antenna elements in vertical and horizontal direction, respectively. Similarly, we can also obtain 𝐚r(ϕi,lr,θi,lr)\mathbf{a}_{r}(\phi^{r}_{i,l},\theta^{r}_{i,l}) with the same form.

The achievable rate of the mmWave MIMO heavily depends on the transmit hybrid precoder, which can be obtained by solving the following optimization problem [12]

max𝐅RF,𝐅BB\displaystyle\max\limits_{\mathbf{F}_{\rm RF},\mathbf{F}_{\rm BB}} k=1Klog2|𝐈+1σn2𝐇[k]𝐅RF𝐅BB[k]𝐅BBH[k]𝐅RFH𝐇H[k]|\displaystyle\sum\nolimits_{k=1}^{K}\!\!\!\!\!\log_{2}|\mathbf{I}\!+\!\tfrac{1}{\sigma_{n}^{2}}\mathbf{H}[k]\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]\mathbf{F}_{\rm BB}^{H}[k]\mathbf{F}_{\rm RF}^{H}\mathbf{H}^{H}[k]| (3)
s.t.\displaystyle\text{s.t. } 𝐅RFRF,k=1K||𝐅RF𝐅BB[k]||F2=KNs,\displaystyle\mathbf{F}_{\rm RF}\in\mathcal{F}_{\rm RF},\sum\nolimits_{k=1}^{K}||\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]||_{F}^{2}=KN_{s},

where RF\mathcal{F}_{\rm RF} is a set of feasible RF precoder satisfying constant-modulus constraint. The coupling between 𝐅RF\mathbf{F}_{\rm RF} and {𝐅BB[k]}k=1K\{\mathbf{F}_{\rm BB}[k]\}_{k=1}^{K} and the constant-modulus constraint of RF\mathcal{F}_{\rm RF} lead to the challenging hybrid precoder design.

III PCA-Based Hybrid Precoder Design for FS

In this section, we derive hybrid precoders for the PCS, in which only a subset of antennas are connected to each RF chain. Our goal is to design the optimal frequency-flat RF precoder from the fully-digital frequency-selective precoder.

III-A Digital Precoder Design

We first design the digital precoder by fixing the RF precoder. Solving (3) can be difficult due to the coupling of the baseband and RF precoders[12]. Therefore, considering 𝐅~BB[k]=(𝐅RFH𝐅RF)12𝐅BB[k]\mathbf{\widetilde{F}}_{\rm BB}[k]=(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{\frac{1}{2}}\mathbf{F}_{\rm BB}[k] to be the equivalent baseband precoder, the equivalent problem of (3) can be expressed as follows

max𝐅RF,𝐅~BB\displaystyle\max\limits_{\mathbf{F}_{\rm RF},\mathbf{\widetilde{F}}_{\rm BB}} k=1Klog2|𝐈+1σn2𝐇[k]𝐅RF(𝐅RFH𝐅RF)12𝐅~BB[k]\displaystyle\sum\nolimits_{k=1}^{K}\!\!\!\!\!\log_{2}|\mathbf{I}\!+\!\tfrac{1}{\sigma_{\rm n}^{2}}\mathbf{H}[k]\mathbf{F}_{\rm RF}(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}\mathbf{\widetilde{F}}_{\rm BB}[k] (4)
×𝐅~BBH[k](𝐅RFH𝐅RF)12𝐅RFH𝐇H[k]|\displaystyle\times\mathbf{\widetilde{F}}_{\rm BB}^{H}[k](\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}\mathbf{F}_{\rm RF}^{H}\mathbf{H}^{H}[k]|
s.t.\displaystyle\text{s.t. } 𝐅RFRF,k=1K||𝐅~BB[k]||F2=KNs.\displaystyle\mathbf{F}_{\rm RF}\in\mathcal{F}_{\rm RF},\sum\nolimits_{k=1}^{K}||\mathbf{\widetilde{F}}_{\rm BB}[k]||_{F}^{2}=KN_{s}.

For the optimization problem (4), we first consider the optimal solution of {𝐅~BB[k]}k=1K\{\mathbf{\widetilde{F}}_{\rm BB}[k]\}_{k=1}^{K}. Specifically, consider the singular value decomposition (SVD) of 𝐇[k]\mathbf{H}[k] associated with the kkth subcarrier as 𝐇[k]=𝐔[k]𝚺[k]𝐕H[k]\mathbf{H}[k]=\mathbf{U}[k]\mathbf{\Sigma}[k]\mathbf{V}^{H}[k], and the SVD of the matrix 𝚺[k]𝐕H[k]𝐅RF(𝐅RFH𝐅RF)1/2=𝐔~[k]𝚺~[k]𝐕~H[k]\mathbf{\Sigma}[k]\mathbf{V}^{H}[k]\mathbf{F}_{\rm RF}(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-1/2}=\mathbf{\widetilde{U}}[k]\mathbf{\widetilde{\Sigma}}[k]\mathbf{\widetilde{V}}^{H}[k]. Therefore, the optimal 𝐅~BB[k]=[𝐕~[k]]:,1:Ns𝚲[k]\mathbf{\widetilde{F}}_{\rm BB}[k]=[\mathbf{\widetilde{V}}[k]]_{:,1:N_{s}}\mathbf{\Lambda}[k], and thus the optimal baseband precoder 𝐅BB[k]\mathbf{F}_{\rm BB}[k] can be expressed as

𝐅BB[k]=\displaystyle\mathbf{F}_{\rm BB}[k]= (𝐅RFH𝐅RF)12𝐅~BB[k]\displaystyle(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}\mathbf{\widetilde{F}}_{\rm BB}[k] (5)
=\displaystyle= (𝐅RFH𝐅RF)12[𝐕~[k]]:,1:Ns𝚲[k],\displaystyle(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}[\mathbf{\widetilde{V}}[k]]_{:,1:N_{s}}\mathbf{\Lambda}[k],

where 𝚲[k]=(μNs/[𝚺~[k]]i,i2)+\mathbf{\Lambda}[k]=(\mu-N_{s}/[\mathbf{\widetilde{\Sigma}}[k]]_{i,i}^{2})^{+} (1iNs1\leq i\leq N_{s}, 1kK1\leq k\leq K) is a water-filling solution matrix, in which μ\mu satisfies k=1Ki=1Ns(μNs/[𝚺~[k]]i,i2)+=KNs\sum_{k=1}^{K}\sum_{i=1}^{N_{s}}(\mu-N_{s}/[\mathbf{\widetilde{\Sigma}}[k]]_{i,i}^{2})^{+}=KN_{s}. Then the problem reduces to obtain the optimal solution of 𝐅RF\mathbf{F}_{\rm RF} to (4).

III-B PCA-Based Precoder Design

Regarding the transmit hybrid precoder for FS, there are NtN_{t} antennas and NtRFN_{t}^{\rm RF} RF chains. For simplicity, we consider the numbers of antennas for different RF chains are identical, and the cardinality of each subset for every antenna group is Ntsub=Nt/NtRFN_{t}^{\rm sub}=N_{t}/N_{t}^{\rm RF}. We define the set of antenna indexes as {1,,Nt}\{1,\cdots,N_{t}\}, and 𝒮r\mathcal{S}_{r} as the subset of the antennas associated with the rrth RF chain, where 𝒮r={(r1)Ntsub+1,,rNtsub}\mathcal{S}_{r}=\{(r-1)N_{t}^{\rm sub}+1,\cdots,rN_{t}^{\rm sub}\}, for 1rNtRF1\leq r\leq N_{t}^{\rm RF}. For the FS, the analog precoder 𝐅RF\mathbf{F}_{\rm RF} can be written as a block diagonal matrix 𝐅RF=blkdiag(𝐟RF,𝒮1,,𝐟RF,𝒮NtRF)\mathbf{F}_{\rm RF}=\text{blkdiag}(\mathbf{f}_{{\rm RF},\mathcal{S}_{1}},\cdots,\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}), where 𝐟RF,𝒮rNtsub×1\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}\in\mathbb{C}^{N_{t}^{\rm sub}\times 1} is the analog beamforming vector associated with the rrth subarray for the rrth RF chain. Defining optimal digital precoder 𝐅opt[k]=[𝐕[k]]:,1:Ns\mathbf{F}_{\rm opt}[k]=[\mathbf{V}[k]]_{:,1:N_{s}} for 1kK1\leq k\leq K, the optimal digital precoder can be expressed as

𝐅optH[k]=[𝐅opt,𝒮1H[k]𝐅opt,𝒮NtRFH[k]],\mathbf{F}_{\rm opt}^{H}[k]=\begin{bmatrix}\mathbf{F}_{{\rm opt},\mathcal{S}_{1}}^{H}[k]\ \cdots\ \mathbf{F}_{{\rm opt},\mathcal{S}_{N_{t}^{\rm RF}}}^{H}[k]\end{bmatrix}, (6)

where 𝐅opt,𝒮r[k]Ntsub×NtRF\mathbf{F}_{{\rm opt},\mathcal{S}_{r}}[k]\in\mathbb{C}^{N_{t}^{\rm sub}\times N_{t}^{\rm RF}}. Moreover, we regard the matrix 𝐅𝒮r=[𝐅opt,𝒮r[1]𝐅opt,𝒮r[2]𝐅opt,𝒮r[K]]\mathbf{F}_{\mathcal{S}_{r}}=\begin{bmatrix}\mathbf{F}_{{\rm opt},\mathcal{S}_{r}}[1]\ \mathbf{F}_{{\rm opt},\mathcal{S}_{r}}[2]\ \cdots\ \mathbf{F}_{{\rm opt},\mathcal{S}_{r}}[K]\end{bmatrix} consisting of the optimal precoder of all subcarriers in the rrth subarray as the data set in the PCA problem [16]. Additionally, to achieve the stable solution with low complexity for PCA, SVD is applied to the data set matrix 𝐅𝒮r\mathbf{F}_{\mathcal{S}_{r}}. This process is detailed in Proposition 1, where its optimality is also verified as follows.

Proposition 1.

For FS, considering 𝐅𝒮r=[𝐅opt,𝒮r[1]𝐅opt,𝒮r[K]]\mathbf{F}_{\mathcal{S}_{r}}=\begin{bmatrix}\mathbf{F}_{{\rm opt},\mathcal{S}_{r}}[1]&\cdots&\mathbf{F}_{{\rm opt},\mathcal{S}_{r}}[K]\end{bmatrix}, the RF precoder 𝐅RF\mathbf{F}_{\rm RF} solving problem (10) with the subarray analog/digital architecture is given by 𝐅=blkdiag(𝐟RF,𝒮1,,𝐟RF,𝒮NtRF)\mathbf{F}=\text{\rm blkdiag}(\mathbf{f}_{{\rm RF},\mathcal{S}_{1}},\cdots,\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}), with 𝐟RF,𝒮r=αr𝐮𝒮r\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}=\alpha_{r}\mathbf{u}_{\mathcal{S}_{r}}, for r=1,,NtRFr=1,\cdots,N_{t}^{\rm RF}, where αr\alpha_{r}\in\mathbb{C} and 𝐮𝒮r\mathbf{u}_{\mathcal{S}_{r}} is the right singular vector corresponding with the largest singular value of the matrix 𝐅𝒮r\mathbf{F}_{\mathcal{S}_{r}}.

Proof.

Following the similar steps of the equations (12)-(14) in [11] and defining [𝚺[k]]1:Ns,1:Ns=𝚺1[k][\mathbf{\Sigma}[k]]_{1:N_{s},1:N_{s}}=\mathbf{\Sigma}_{1}[k], the objective function in problem (3) can be approximate as

k=1Klog2|𝐈+1σn2𝐇[k]𝐅RF𝐅BB[k]𝐅BBH[k]𝐅RFH𝐇H[k]|\displaystyle\sum\nolimits_{k=1}^{K}\log_{2}|\mathbf{I}+\tfrac{1}{\sigma_{n}^{2}}\mathbf{H}[k]\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]\mathbf{F}_{\rm BB}^{H}[k]\mathbf{F}_{\rm RF}^{H}\mathbf{H}^{H}[k]| (7)
\displaystyle\approx k=1K(log2|𝐈Ns+1σn2𝚺12[k]|(Ns𝐅optH[k]𝐅RF𝐅BB[k]F2)).\displaystyle\sum\nolimits_{k=1}^{K}\!(\log_{2}\!|\mathbf{I}_{N_{s}}\!\!+\!\!\tfrac{1}{\sigma_{n}^{2}}\!\mathbf{\Sigma}_{1}^{2}[k]|\!-\!(\!N_{s}\!\!-\!\!||\mathbf{F}_{\rm opt}^{H}\![k]\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]||_{F}^{2})\!).

Therefore, the optimization problem (3) is equivalent to the following optimization problem

max𝐅RF,𝐅BB\displaystyle\max\limits_{\mathbf{F}_{\rm RF},\mathbf{F}_{\rm BB}} k=1K𝐅optH[k]𝐅RF𝐅BB[k]F2\displaystyle\sum\nolimits_{k=1}^{K}||\mathbf{F}_{\rm opt}^{H}[k]\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]||_{F}^{2} (8)
s.t.\displaystyle\text{s.t. } 𝐅RFRF,k=1K||𝐅RF𝐅BB[k]||F2=KNs,\displaystyle\mathbf{F}_{\rm RF}\in\mathcal{F}_{\rm RF},\sum\nolimits_{k=1}^{K}||\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]||_{F}^{2}=KN_{s},

where RF\mathcal{F}_{\rm RF} is a set of feasible RF precoder satisfying constant-modulus constraint. The objective function in (8) is

k=1K\displaystyle\sum\nolimits_{k=1}^{K} 𝐅optH[k]𝐅RF𝐅BB[k]F2=k=1KTr(𝐅optH[k]𝐅RF(𝐅RFH𝐅RF)12CLOSE\displaystyle||\mathbf{F}_{\rm opt}^{H}\![k]\mathbf{F}_{\rm RF}\mathbf{F}_{\rm BB}[k]||_{F}^{2}\!=\!\!\sum\nolimits_{k=1}^{K}\!\!\!\!\!\!\!\!\text{Tr}(\mathbf{F}_{\rm opt}^{H}\![k]\mathbf{F}_{\rm RF}\!(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF}\!)\!^{-\!\frac{1}{2}} (9)
×(𝐅~BB[k]𝐅~BBH[k])(𝐅RFH𝐅RF)12𝐅RFH𝐅opt[k]).\displaystyle\times(\mathbf{\widetilde{F}}_{\rm BB}[k]\mathbf{\widetilde{F}}_{\rm BB}^{H}[k])(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm opt}[k]).

According to previous work [17], unitary constraints offer a close performance to the total power constraint while providing a relatively simple form of solution. To simplify the problem, we consider condition under unitary power constraints instead. Therefore, water-filling power allocation coefficients can be ignored. In detail, the equivalent baseband precoder 𝐅~BB[k]=[𝐕~[k]]:,1:Ns\mathbf{\widetilde{F}}_{\rm BB}[k]=[\mathbf{\widetilde{V}}[k]]_{:,1:N_{s}}, which means that 𝐅~BB[k]\mathbf{\widetilde{F}}_{\rm BB}[k] is a unitary or simi-unitary matrix depending on the relationship between NsN_{s} and NtRFN_{t}^{\rm RF}. When Ns=NtRFN_{s}=N_{t}^{\rm RF}, 𝐅~BB[k]𝐅~BBH[k]\mathbf{\widetilde{F}}_{\rm BB}[k]\mathbf{\widetilde{F}}_{\rm BB}^{H}[k] is 𝐈Ns\mathbf{I}_{N_{s}}. When Ns<NtRFN_{s}<N_{t}^{\rm RF}, denoting the SVD of 𝐅~BB[k]=𝐔BB[k][𝐈𝐍𝐬 0]T𝐕BBH[k]\mathbf{\widetilde{F}}_{\rm BB}[k]=\mathbf{U}_{\rm BB}[k]\begin{bmatrix}\mathbf{I_{N_{s}}}\ \mathbf{0}\end{bmatrix}^{T}\mathbf{V}_{\rm BB}^{H}[k], thus 𝐅~BB[k]𝐅~BBH[k]=𝐔BB[k]blkdiag(𝐈Ns,𝟎NtRFNs)𝐔BBH[k]\mathbf{\widetilde{F}}_{\rm BB}[k]\mathbf{\widetilde{F}}_{\rm BB}^{H}[k]=\mathbf{U}_{\rm BB}[k]\text{blkdiag}(\mathbf{I}_{N_{s}},\mathbf{0}_{N_{t}^{\rm RF}-N_{s}})\mathbf{U}_{\rm BB}^{H}[k]. Therefore, the solution to the condition when Ns=NtRFN_{s}=N_{t}^{\rm RF} will also suffice the condition when Ns<NtRFN_{s}<N_{t}^{\rm RF}. Therefore, the objective function of (8) goes down to

Tr(𝐅optH[k]𝐅RF(𝐅RFH𝐅RF)12(𝐅RFH𝐅RF)12𝐅RFH𝐅opt[k])\displaystyle\text{Tr}(\mathbf{F}_{\rm opt}^{H}[k]\mathbf{F}_{\rm RF}(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm opt}[k]) (10)
=\displaystyle= 𝐅optH[k]𝐅RF(𝐅RFH𝐅RF)12F2.\displaystyle||\mathbf{F}_{\rm opt}^{H}[k]\mathbf{F}_{\rm RF}(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})^{-\frac{1}{2}}||_{F}^{2}.

For simplicity, we denote 𝐅RF(𝐅RFH𝐅RF)1/2\mathbf{F}_{\rm RF}{(\mathbf{F}_{\rm RF}^{H}\mathbf{F}_{\rm RF})}^{-1/2} as 𝐅¯RF\mathbf{\bar{F}}_{\rm RF}. Therefore, 𝐅¯RF\mathbf{\bar{F}}_{\rm RF} can be written into following block diagram matrix

𝐅¯RF=blkdiag(𝐟RF,𝒮1|𝐟RF,𝒮1|1,,𝐟RF,𝒮NtRF|𝐟RF,𝒮NtRF|1).\mathbf{\bar{F}}_{\rm RF}\!=\!\text{blkdiag}(\mathbf{f}_{{\rm RF},\mathcal{S}_{1}}|\mathbf{f}_{{\rm RF},\mathcal{S}_{1}}|^{-1}\!,\cdots,\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}|\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}|^{-1}). (11)

By substituting (11) and (6) into (10), the objective function of the optimization problem can be further expressed as

k=1K𝐅optH[k]𝐅¯RFF2\displaystyle\sum\nolimits_{k=1}^{K}||\mathbf{F}_{\rm opt}^{H}[k]\mathbf{\bar{F}}_{\rm RF}||_{F}^{2} (12)
=\displaystyle= k=1K[𝐟RF,𝒮1𝐅opt,𝒮1H[k]|𝐟RF,𝒮1|𝐟RF,𝒮NtRF𝐅opt,𝒮NtRFH[k]|𝐟RF,𝒮NtRF|]F2\displaystyle\sum_{k=1}^{K}||\begin{bmatrix}\frac{\mathbf{f}_{{\rm RF},\mathcal{S}_{1}}\mathbf{F}_{{\rm opt},\mathcal{S}_{1}}^{H}[k]}{|\mathbf{f}_{{\rm RF},\mathcal{S}_{1}}|}&\cdots&\frac{\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}\mathbf{F}_{{\rm opt},\mathcal{S}_{N_{t}^{\rm RF}}}^{H}[k]}{|\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}|}\end{bmatrix}||_{F}^{2}
=\displaystyle= r=1NtRF𝐟RF,𝒮r𝐅𝒮rH𝐅𝒮r𝐟RF,𝒮rH|𝐟RF,𝒮r|2.\displaystyle\sum_{r=1}^{N_{t}^{\rm RF}}\frac{\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}\mathbf{F}_{\mathcal{S}_{r}}^{H}\mathbf{F}_{\mathcal{S}_{r}}\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}^{H}}{|\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}|^{2}}.

Therefore, the solution to the optimization problem (10) is maxk=1K𝐅RF𝐅¯RF𝐅optH[k]F2=r=1NtRFλ12(𝐅𝒮r)\max_{\mathbf{F}_{\rm RF}}\sum_{k=1}^{K}||\mathbf{\bar{F}}_{\rm RF}\mathbf{F}_{\rm opt}^{H}[k]||_{F}^{2}=\sum_{r=1}^{N_{t}^{\rm RF}}\lambda_{1}^{2}(\mathbf{F}_{\mathcal{S}_{r}}). The maximum value can only be obtained when 𝐟RF,𝒮r=αr𝐮𝒮r,1\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}=\alpha_{r}\mathbf{u}_{\mathcal{S}_{r},1}, where αr\alpha_{r} is an arbitrary complex value, and 𝐮𝒮r\mathbf{u}_{\mathcal{S}_{r}} is the largest singular value of the matrix 𝐅𝒮r\mathbf{F}_{\mathcal{S}_{r}}. ∎

Taking the constraint of RF precoder into account, we can design the RF precoder by solving

𝐅RF\displaystyle\mathbf{F}_{\rm RF} =blkdiag(𝐟RF,𝒮1,,𝐟RF,𝒮NtRF)\displaystyle=\text{blkdiag}(\mathbf{f}_{{\rm RF},\mathcal{S}_{1}},\cdots,\mathbf{f}_{{\rm RF},\mathcal{S}_{N_{t}^{\rm RF}}}) (13)
where,\displaystyle\text{where, } 𝐟RF,𝒮r=argmin𝐱,|[𝐱]i,j|=1/Nsub𝐱𝐮𝒮rF2,\displaystyle\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}=\arg\min_{\mathbf{x},|[\mathbf{x}]_{i,j}|=1/\sqrt{N_{\rm sub}}}||\mathbf{x}-\mathbf{u}_{\mathcal{S}_{r}}||^{2}_{F},
for r=1,,NtRF.\displaystyle\text{for }r=1,\cdots,N_{t}^{\rm RF}.

With the constant-modulus constraint, the set of possible 𝐟RF,𝒮r\mathbf{f}_{{\rm RF},\mathcal{S}_{r}} is actually a hypersphere in the space of Nt×1\mathbb{C}^{N_{t}\times 1}, and 𝐮𝒮r\mathbf{u}_{\mathcal{S}_{r}} is a known point in the space of Nt×1\mathbb{C}^{N_{t}\times 1}. Therefore, the optimization problem in (13) is actually a distance minimization problem. Therefore, the solution is the point on this hypersphere sharing same direction of the know point [𝐟RF,𝒮r]i=Nsubej([𝐮𝒮r]i)[\mathbf{f}_{{\rm RF},\mathcal{S}_{r}}]_{i}=\sqrt{N_{\rm sub}}e^{j\angle([\mathbf{u}_{\mathcal{S}_{r}}]_{i})}.

When the quantization of phase shifters is considered, we assume the quantization bits are QQ. Therefore, the phase shifters can only be chosen from the following quantized phase set 𝒬={0,2π2Q,,2π(2Q1)2Q}\mathcal{Q}=\{0,\frac{2\pi}{2^{Q}},\cdots,\frac{2\pi(2^{Q}-1)}{2^{Q}}\}. Specifically, after obtaining the RF precoder 𝐅RF\mathbf{F}_{\rm RF}, the quantization process can be realized by searching for the minimum Euclidean distance between ([𝐅RF]i,j)\angle([\mathbf{F}_{\rm RF}]_{i,j}) and quantized phase from 𝒬\mathcal{Q}.

IV Shared-AHC Algorithm for DS Grouping

In Section III, we have found the SE performance of FS heavily depends on {𝐅𝒮r}r=1NtRF\{\mathbf{F}_{\mathcal{S}_{r}}\}_{r=1}^{N_{t}^{\rm RF}}. This observation motivates us to optimize the antenna grouping {𝒮r}r=1NtRF\{\mathcal{S}_{r}\}_{r=1}^{N_{t}^{\rm RF}} to further improve the SE performance when DS is considered.

The DS problem can be formulated as follows

maxr=1NtRF𝒮1,,𝒮NtRFλ12(𝐅𝒮r)\displaystyle\max\limits_{\mathcal{S}_{1},\cdots,\mathcal{S}_{N_{t}^{\rm RF}}}\sum\nolimits_{r=1}^{N_{t}^{\rm RF}}\lambda_{1}^{2}(\mathbf{F}_{\mathcal{S}_{r}}) (14)
s.t. r=1NtRF𝒮r={1,,Nt},𝒮i𝒮j= for ij,𝒮rr.\displaystyle\text{s.t. }\cup_{r=1}^{N_{t}^{\rm RF}}\!\mathcal{S}_{r}\!=\!\{1,\!\cdots\!,\!N_{t}\},\ \mathcal{S}_{i}\!\cap\!\mathcal{S}_{j}\!=\!\emptyset\text{ for }i\!\not=\!j,\ \mathcal{S}_{r}\!\not=\!\emptyset\ \forall r.

This optimization problem is a combinational optimization problem, which requires an exhaustive search to reach the optimal solution. To obtain the optimal solution, the number of all possible combinations for exhaustive search can be 1(NtRF)!n=0NrRF(1)NtRFn(NtRFn)nNt\frac{1}{(N_{t}^{\rm RF})!}\sum_{n=0}^{N_{r}^{\rm RF}}(-1)^{N_{t}^{\rm RF}-n}\binom{N_{t}^{\rm RF}}{n}n^{N_{t}} according to [18], which is a very large number. To illustrate, when Nt=64N_{t}=64 and NtRF=4N_{t}^{\rm RF}=4, the number of all possible combinations can be up to 1.4178×10371.4178\times 10^{37}.

Therefore, a low-complexity algorithm need to be develop to solve problem (14). Specifically, we use the Minkowski 2\ell_{2}-norm [19] to estimate the square of the singular value of the matrix 𝐅𝒮r\mathbf{F}_{\mathcal{S}_{r}} by λ12(𝐅𝒮r)=λ1(𝐑𝒮r)1|𝒮r|i=1|𝒮r|j=1|𝒮r||[𝐑𝒮r]i,j|=1|𝒮r|i𝒮rj𝒮r|[𝐑F]i,j|\lambda_{1}^{2}(\mathbf{F}_{\mathcal{S}_{r}})=\lambda_{1}(\mathbf{R}_{\mathcal{S}_{r}})\approx\frac{1}{|\mathcal{S}_{r}|}\sum_{i=1}^{|\mathcal{S}_{r}|}\sum_{j=1}^{|\mathcal{S}_{r}|}|[\mathbf{R}_{\mathcal{S}_{r}}]_{i,j}|=\frac{1}{|\mathcal{S}_{r}|}\sum_{i\in\mathcal{S}_{r}}\sum_{j\in\mathcal{S}_{r}}|[\mathbf{R}_{F}]_{i,j}|, where 𝐑𝒮r=𝐅𝒮r𝐅𝒮rH\mathbf{R}_{\mathcal{S}_{r}}=\mathbf{F}_{\mathcal{S}_{r}}\mathbf{F}_{\mathcal{S}_{r}}^{H} and 𝐑F=𝐅𝐅H\mathbf{R}_{F}=\mathbf{F}\mathbf{F}^{H}.

To reduce the complexity while achieve the good SE performance, we consider the antenna grouping from the viewpoint of clustering analysis in machine learning. To be specific, we propose a shared-AHC algorithm as listed in Algorithm 1, which is developed from the AHC algorithm in machine learning to group the antennas into different subarrays associated with different RF chains. Traditional AHC algorithm is a clustering algorithm that builds a cluster hierarchy from the bottom up. It starts by adding all data to multiple clusters, followed by iteratively pair-wise merging these clusters until only one cluster is left at the top of the hierarchy. The shared-AHC algorithm is different from the traditional AHC algorithm [15] in two distinguished aspects. First, the aim of clustering in our antenna grouping problem is to build NtRFN_{t}^{\rm RF} clusters instead of only one cluster in conventional AHC algorithm. Second, the pair-wise merging criterion in the proposed algorithm is “shared”, while the conventional AHC algorithm only considers the target cluster. To further illustrate the “shared” mechanism, we introduce the metric of mutual correlation g(𝒮n,𝒮m)g(\mathcal{S}_{n},\mathcal{S}_{m}) between the cluster 𝒮n\mathcal{S}_{n} and 𝒮m\mathcal{S}_{m}

g(𝒮n,𝒮m)=1|𝒮n||𝒮m|i𝒮nj𝒮m|[𝐑F]i,j|.g(\mathcal{S}_{n},\mathcal{S}_{m})=\frac{1}{|\mathcal{S}_{n}||\mathcal{S}_{m}|}\sum_{i\in\mathcal{S}_{n}}\sum_{j\in\mathcal{S}_{m}}|[\mathbf{R}_{F}]_{i,j}|. (15)

In each clustering iteration, we first focus on a cluster 𝒮n\mathcal{S}_{n}, and find a cluster 𝒮m\mathcal{S}_{m} maximizes g(𝒮n,𝒮l)g(\mathcal{S}_{n},\mathcal{S}_{l}) among all possible 𝒮l\mathcal{S}_{l}. If the cluster 𝒮n\mathcal{S}_{n} also maximizes g(𝒮m,𝒮l)g(\mathcal{S}_{m},\mathcal{S}_{l}) among all possible 𝒮l\mathcal{S}_{l}, we merge 𝒮n\mathcal{S}_{n} and 𝒮m\mathcal{S}_{m}. Otherwise, the cluster 𝒮n\mathcal{S}_{n} and cluster 𝒮m\mathcal{S}_{m} are not merged and algorithm goes into the next iteration. Therefore, our proposed algorithm is featured as “shared”, since two clusters mutually share the maximum correlation in the sense of (15). This process is realized in Algorithm 1.

Algorithm 1 Shared Agglomerative Hierarchical Clustering (Shared-AHC) Algorithm for DS Grouping.
1: 𝐑F\mathbf{R}_{F}, number of antennas and RF chains NtN_{t}, NtRFN_{t}^{\rm RF}.
2: Grouping result 𝒮1,,𝒮NtRF\mathcal{S}_{1},\cdots,\mathcal{S}_{N_{t}^{\rm RF}}.
3: Nsub=NtN_{\rm sub}=N_{t}, 𝒮i={i}\mathcal{S}_{i}=\{i\} for i=1,,Nti=1,\cdots,N_{t}
4: while Nsub>NtRFN_{\rm sub}>N_{t}^{\rm RF} do
5:   𝒮i0=𝒮i\mathcal{S}_{i}^{0}=\mathcal{S}_{i} for i=1,,Nsubi=1,\cdots,N_{\rm sub}, nsub=1n_{\rm sub}=1
6:   for i=1:Nsubi=1:N_{\rm sub} do
7:    if r0 s.t. 𝒮i0𝒮r0\exists r_{0}\text{ s.t. }\mathcal{S}_{i}^{0}\in\mathcal{S}_{r_{0}} then continue
8:    else if i=Nsubi=N_{\rm sub} then 𝒮nsub=𝒮i0\mathcal{S}_{n_{\rm sub}}=\mathcal{S}_{i}^{0}
9:    else
10:      j=argmaxl{i+1,,Nsub}g(𝒮i,𝒮l)j=\arg\max\limits_{l\in\{i+1,\cdots,N_{\rm sub}\}}g(\mathcal{S}_{i},\mathcal{S}_{l})
11:      i0=argmaxl{1,,Nsub}{j}g(𝒮j,𝒮l)i^{0}=\arg\max\limits_{l\in\{1,\cdots,N_{\rm sub}\}\setminus\{j\}}g(\mathcal{S}_{j},\mathcal{S}_{l})
12:      if i=i0i=i^{0} then 𝒮nsub=𝒮i𝒮j\mathcal{S}_{n_{\rm sub}}=\mathcal{S}_{i}\cup\mathcal{S}_{j}
13:      else 𝒮nsub=𝒮i\mathcal{S}_{n_{\rm sub}}=\mathcal{S}_{i}
14:      end if
15:    end if
16:    nsub=nsub+1n_{\rm sub}=n_{\rm sub}+1
17:   end for
18:   Nsub0=nsub1N_{\rm sub}^{0}=n_{\rm sub}-1
19:   if Nsub0<NrRFN_{\rm sub}^{0}<N_{r}^{\rm RF} then 𝒮i=𝒮i0\mathcal{S}_{i}=\mathcal{S}_{i}^{0} for i=1,,Nsubi=1,\cdots,N_{\rm sub}
20:     break
21:   else Nsub=Nsub0N_{\rm sub}=N_{\rm sub}^{0}
22:   end if
23: end while
24: if Nsub>NtRFN_{\rm sub}>N_{t}^{\rm RF} then
25:   Sort 𝒮i\mathcal{S}_{i} according to the ascending order of cardinality
26:   for i=1:NtRFNsubi=1:N_{t}^{\rm RF}-N_{\rm sub} do
27:    j=argmaxl={NtRFNsub+1,,Nsub}g(𝒮i,𝒮l)j=\arg\max\limits_{l=\{N_{t}^{\rm RF}-N_{\rm sub}+1,\cdots,N_{\rm sub}\}}g(\mathcal{S}_{i},\mathcal{S}_{l})
28:    𝒮i=𝒮i𝒮j\mathcal{S}_{i}=\mathcal{S}_{i}\cup\mathcal{S}_{j}
29:   end for
30:   Rearrange the subscript to guarantee that the order of subscripts are from 1 to NtRFN_{t}^{\rm RF}
31: end if

V Energy Efficiency Analysis

The implementation of PCS not only reduces the hardware complexity, but also improves the EE. In this section, we analyze the EE of the designs. Define the EE as η=RB/P\eta=RB/P, where BB is the bandwidth of the channel, and PP is the total power consumption of the system.

Different connection patterns between the phase shifters and antennas can influence the power consumption. Because the number of phase shifters is different in different connection patterns. In this system, FCA use up to NtNtRFN_{t}N_{t}^{\rm RF} phase shifter for each RF chain connecting to every antennas. While the PCS use NtN_{t} phase shifters.

Furthermore, the different antenna architectures should also be taken into account regarding the total power. Specifically, we consider the hybrid MIMO system using passive antennas and active antennas as shown in Fig. 10 of [20]. Both of them consist electronic components such as digital-analog convertors (DAC), power amplifiers (PA), local oscillators (LO), and mixers etc. The main difference between active and passive antenna architecture lies in the number of the PAs. In passive antennas, the number of PAs is the same as that of the RF chains. While for active antennas, the number of PAs is the same as that of antennas. This difference can lead to different power consumption because the PAs are heavily power-consuming. Thus we will analyze the power consumption of the two different antenna architectures, respectively.

Given the above antenna architecture, the power consumption for FCA and PCS are respectively PFCAp=NtNtRFPPS+NtRF(PDAC+Pmix+PPA+PLO)P_{\rm FCA}^{p}=N_{t}N_{t}^{\rm RF}P_{\rm PS}\!+\!N_{t}^{\rm RF}(P_{\rm DAC}\!+\!P_{\rm mix}\!+\!P_{\rm PA}\!+\!P_{\rm LO}) and PPCSp=NtPPS+NtRF(PDAC+Pmix+PPA+PLO)P_{\rm PCS}^{p}=N_{t}P_{\rm PS}\!+\!N_{t}^{\rm RF}(P_{\rm DAC}+P_{\rm mix}\!+\!P_{\rm PA}\!+\!P_{\rm LO}). By contrast, the power consumption for FCA and PCS with active antenna architecture are PFCAa=NtNtRFPPS+NtPPA+NtRF(PDAC+Pmix+PLO)P_{\rm FCA}^{a}=N_{t}N_{t}^{\rm RF}P_{\rm PS}\!+\!N_{t}P_{\rm PA}\!+\!N_{t}^{\rm RF}(P_{\rm DAC}\!+\!P_{\rm mix}\!+\!P_{\rm LO}\!) and PPCSa=NtPPS+NtPPA+NtRF(PDAC+Pmix+PLO)P_{\rm PCS}^{a}=N_{t}P_{\rm PS}\!+\!N_{t}P_{\rm PA}\!+\!N_{t}^{\rm RF}(P_{\rm DAC}\!+\!P_{\rm mix}\!+\!P_{\rm LO}). Moreover, according to the antenna architecture for fully-digital (FD), the power consumption is PFD=Nt(PPA+PDAC+Pmix+PLO)P_{\rm FD}=N_{t}(P_{\rm PA}+P_{\rm DAC}+P_{\rm mix}+P_{\rm LO}). Additionally, the power consumption of electronic components in the three architectures are phase shifter PPS=30P_{\rm PS}=30 mW [21], DAC PDAC=200P_{\rm DAC}=200 mW [21], mixer Pmix=39P_{\rm mix}=39 mW [22], LO PLO=5P_{\rm LO}=5 mW [21], and PA PPA=138P_{\rm PA}=138 mW [23].

VI Simulations

In this section, we investigate the SE and EE performance for the hybrid precoder design. For the channel model, we adopt Dirac delta function as the pulse shaping filter and a cyclic prefix with the length of D=64D=64. The number of subcarriers is K=512K=512. The transmission bandwidth is B=500B=500 MHz. We consider that the path delay is uniformly distributed in [0,DTs][0,DT_{s}] (Ts=1/BT_{s}=1/B is the symbol period). The number of the clusters is Ncl=8N_{\rm cl}=8, and azimuth/elevation AoAs and AoDs follow the uniform distribution 𝒰[π/2,π/2]\mathcal{U}[-\pi/2,\pi/2] with angle spread of 7.57.5^{\circ}. Within each cluster, there are Nray=10N_{\rm ray}=10 rays. As for the antennas, we consider transmitter adopt 8×88\times 8 UPA with hybrid precoder, the receiver adopt 2×22\times 2 UPA with fully-digital combiner, and the distance between each adjacent antennas is half wavelength. Moreover, we consider the number of RF chains at transmitter is NtRF=4N_{t}^{\rm RF}=4 and the data stream is Ns=3N_{s}=3. Additionally, we consider 4 types of classical FS patterns shown in Fig. 1, where antenna elements with the same color share the same RF chain.

Refer to caption
Fig. 1: Four types of FS: (a) Vertical type; (b) Horizontal type; (c) Squared type; (d) Interlaced type.

Throughout this part, following baselines will be considered for performance benchmarks: Optimal fully-digital scheme considers the fully-digital MIMO system, where the SVD-based precoder/combiner is adopted as the performance upper bound. Simultaneous OMP (SOMP) scheme is an extension version of the narrow-band OMP-based spatially sparse precoding in [11]. In broadband, SOMP-based hybrid precoding scheme can simultaneously design the RF precoder/combiner for all subcarriers. Discrete Fourier transform (DFT) codebook scheme designs the RF precoder/combiner from the DFT codebook instead of steering vectors codebook in SOMP scheme [9]. Covariance eigenvalue decomposition (EVD) scheme estimates the covariance matrix of the channels using the mean of auto-correlation matrices at each subcarrier [13]. The RF precoder is designed based on the EVD of the covariance matrix of the channels.

Refer to caption
Refer to caption
Fig. 2: SE performance comparison of different hybrid precoder schemes: (a) FCA with Q=Q=\infty and Q=3Q=3; (b) PCS with Q=Q=\infty.

In Fig. 2, we compare the SE performance of the proposed hybrid precoding scheme with the baselines, where both FCA and PCS are investigated. In Fig. 2 (a), for FCA, our proposed PCA-based hybrid precoding scheme outperforms conventional DFT codebook-based hybrid precoding scheme and SOMP-based hybrid precoding scheme. Both the proposed PCA-based hybrid precoding scheme and covariance EVD-based hybrid precoding scheme have the very similar performance, and they suffer from negligible performance loss when compared to the optimal fully-digital scheme. This is because mmWave MIMO channels associated with different subcarriers share the same row space due to the common scatterers. Meanwhile, our proposed algorithm can exploit the principal components of the common row space to establish the hybrid precoder. The SOMP-based and DFT codebook-based hybrid precoding schemes work poorly, since their analog codebooks are limited to the steering vector forms. Finally, it can also be observed that the influence of quantization in phase shifters is negligible for our scheme. As for the PCS, Fig. 2 (b) shows that our scheme outperforms conventional covariance EVD-based hybrid precoding scheme with different FS patterns and DS. The antenna grouping scheme in [13] considers a greedy approach, which may lead to the imbalance antenna grouping by acquiring local optimal solution. By contrast, the shared-AHC algorithm for DS grouping introduces the mutually correlation metric (15), which can efficiently avoid this issue. Therefore, the proposed shared-AHC algorithm for DS grouping outperforms it counterpart in [13].

Refer to caption
Refer to caption
Fig. 3: EE performance comparison of different hybrid precoding schemes on different antenna architectures: (a) Passive antenna; (b) Active antenna with.

In Fig. 3, we compare the EE performance of the proposed PCA-based hybrid precoding scheme and the baselines with FCA and PCS, where both passive and active antenna architectures are investigated. Note that the power values of key electronic components can refer to Section V. In Fig. 3 (a), for passive antenna architecture, the EE performance of PCS by using the proposed PCA-based hybrid precoding scheme outperforms that of FCA by using the SOMP-based and DFT codebook-based hybrid precoding schemes. The reason is that PCS adopts a much smaller number of phase shifters than FCA. Moreover, DS outperforms the other FS patterns in SE, and it consumes the same power with the other FS patterns. Therefore, DS outperforms other four types of FS patterns. It is worth mentioning that the optimal fully-digital scheme has the worst EE performance, since the numbers of power-consuming PAs, DACs, and mixers are proportional to that of antennas. In Fig. 3 (b), for active antenna architecture, the advantage of EE performance for different FS patterns by using the proposed hybrid precoding scheme over the FCA with several typical hybrid precoding schemes and optimal fully-digital precoding scheme is not considerable. This is because active antenna architecture requires the power-hungry PAs for each antenna. Meanwhile, the advantage of the reduced power consumption of FS structure is greatly weakened by its disadvantage in SE performance when compared to FCA. Finally, the EE performance of DS with the proposed hybrid precoding scheme still has the obvious advantage over the baselines and four typical types of FS with the proposed scheme. This reveals the appealing advantage of DS in practical situation when both the power consumption and SE should be well balanced.

VII Conclusions

This paper has proposed a hybrid precoding scheme based on machine learning for broadband mmWave MIMO systems with DS. We first acquire the low-dimensional frequency flat precoder from the optimal frequency-selective precoders based on PCA for FS. Then, we extend the proposed PCA-based hybrid precoder design to the DS. We propose the shared-AHC algorithm inspired by cluster analysis in machine learning for antenna grouping to further improve the SE performance. Additionally, we analyze the EE performance for FCA, FS, and DS with passive and active antennas. Simulations further verify the proposed PCA-based hybrid precoding scheme has the better SE and EE performance than conventional schemes.

Acknowledgment

This work was supported by the National Natural Science Foundation of China (Grant Nos. 61471037, 61701027, and 61201181), the Beijing Natural Science Foundation (Grant No. 4182055), Huawei Innovation Research Program (HIRP), and Youth Project of China Academy of Information and Communications Technology.

References

  • [1] Z. Xiao, P. Xia, and X. G. Xia, “Codebook design for millimeter-wave channel estimation with hybrid precoding structure,” IEEE Trans. Wireless Commun., vol. 16, no. 1, pp. 141-153, Jan. 2017.
  • [2] Y. Sun and C. Qi, “Weighted sum-rate maximization for analog beamforming and combining in millimeter wave massive MIMO communications,” IEEE Wireless Commun. Lett., vol. 21, no. 8, pp. 1883-1886, Oct. 2017
  • [3] Z. Gao et al., “MmWave massive-MIMO-based wireless backhaul for the 5G ultra-dense network,” IEEE Wireless Commun., vol. 22, no. 5, pp. 13-21, Oct. 2015.
  • [4] Z. Gao et al., “Compressive sensing techniques for next-generation wireless communications,” IEEE Wireless Commun., vol. 25, no. 3, pp. 144-153, Jun. 2018.
  • [5] A. Liao et al., “2D unitary ESPRIT based super-resolution channel estimation for millimeter-wave massive MIMO with hybrid precoding,” IEEE Access, vol. 5, pp. 24747-24757, 2017.
  • [6] S. He et al., “Codebook-based hybrid precoding for millimeter wave multiuser systems,” IEEE Trans. Signal Process., vol. 65, no. 20, pp. 5289-5304, Oct. 2017.
  • [7] A. Liu and V. K. N. Lau, “Impact of CSI knowledge on the codebook-based hybrid beamforming in massive MIMO,” IEEE Trans. Signal Process., vol. 64, no. 24, pp. 6545-6556, Dec. 2016.
  • [8] S. He, C. Qi, Y. Wu, and Y. Huang, “Energy-efficient transceiver design for hybrid sub-array architecture MIMO systems,” IEEE Access, vol. 4, pp. 9895-9905, 2016.
  • [9] J. Mao et al., “Over-sampling codebook-based hybrid minimum sum-mean-square-error precoding for millimeter-wave 3D-MIMO,” IEEE Wireless Commun. Lett., vol. PP, no. PP, pp. 1-1, May 2018.
  • [10] Y. Huang, J. Zhang, and M. Xiao, “Constant envelope hybrid precoding for directional millimeter-wave communications,” IEEE J. Sel. Areas Commun., vol. PP, no. PP, pp. 1-1, Apr. 2018.
  • [11] O. E. Ayach et al., “Spatially sparse precoding in millimeter wave MIMO systems,” IEEE Trans. Wireless Commun., vol. 13, no. 3, pp. 1499-1513, Mar. 2014.
  • [12] A. Alkhateeb and R. W. Heath Jr., “Frequency selective hybrid precoding for limited feedback millimeter wave systems,” IEEE Trans. Commun., vol. 64, no. 5, pp. 1801-1818, May 2016.
  • [13] S. Park, A. Alkhateeb, and R. W. Heath Jr., “Dynamic subarrays for hybrid precoding in wideband mmWave MIMO system,” IEEE Trans. Wireless Commun., vol. 16, no. 5, pp 2907-2920, May 2017.
  • [14] K. Venugopal, N. G. Prelcic, and R. W. Heath Jr., “Optimality of frequency flat precoding in frequency selective millimeter wave channels,” IEEE Wireless Commun. Lett., vol. 6, no. 3, pp. 330-333, Jun. 2017.
  • [15] S. Zhou, Z. Xu, and F. Liu, “Method for determining the optimal number of clusters based on agglomerative hierarchical clustering,” IEEE Trans. Neural Netw. Learn. Syst., vol. 28, no. 12, pp. 3007-3017, Dec. 2017.
  • [16] C. M. Bishop, Pattern Recognition and Machine Learning. New York, NY, USA: Springer, 2006.
  • [17] D. J. Love et al., “An overview of limited feedback in wireless communication systems,” IEEE J. Sel. Areas Commun., vol. 26, no. 8, pp. 1341-1365, Oct. 2008.
  • [18] R. Graham, D. Knuth, and O. Patashnik, Concrete Mathematics. Reading, MA, USA: Addison-Wesley, 1988.
  • [19] H. Lütkepohl, Handbook Matrics. Hoboken, NJ, USA: Wiley, 1996.
  • [20] W. Hong et al., “Multibeam antenna technologies for 5G wireless communications,” IEEE Trans. Antennas Propag., vol. 65, no. 12, pp. 6231-6249, Dec. 2017.
  • [21] R. Méndez-Rail et al., “Hybrid MIMO architectures for millimeter wave communications: phase shifters or switches?”, IEEE Access, vol. 4, pp. 247-267, Jan. 2016.
  • [22] M. Kraemer, D. Dragomirescu, and R. Plana, “Design of a very low-power, low-cost 60 GHz receiver front-end implemented in 65 nm CMOS technology,” Int. J. Microw. Wireless Technol., vol. 3, pp. 131-138, Apr. 2011.
  • [23] Y. Yu et al., “A 60 GHz phase shifter integrated with LNA and PA in 65 nm CMOS for phased array systems,” IEEE J. Solid-State Circuits, vol. 45, no. 9, pp. 1697-1709, Sep. 2010.