If I have independent variables [x1, x2, x3] If I fit linear regression in sklearn it will give me something like this: <pre class="prettyprint"><code>y = a*x1 + b*x2 + c*x3 + intercept </code></pre> Polynomial regression with poly =2 will give me something like <pre class="prettyprint"><code>y = a*x1^2 + b*x1*x2 ...... </code></pre> I don't want to have terms with second degree like x1^2. how can I get <pre class="prettyprint"><code>y = a*x1 + b*x2 + c*x3 + d*x1*x2 </code></pre> if x1 and x2 have high correlation larger than some threshold value j .

Use patsy to construct a design matrix as follows: <pre class="prettyprint"><code>y, X = dmatrices('y ~ x1 + x2 + x3 + x1:x2', your_data) </code></pre> Where <code>your_data</code> is e.g. a DataFrame with response column <code>y</code> and input columns <code>x1</code>, <code>x2</code> and <code>x3</code>. Then just call the <code>fit</code> method of your estimator, e.g. <code>LinearRegression().fit(X,y)</code>.

How to add interaction term in Python sklearn

Tags:

If I have independent variables [x1, x2, x3] If I fit linear regression in sklearn it will give me something like this:

y = a*x1 + b*x2 + c*x3 + intercept

Polynomial regression with poly =2 will give me something like

y = a*x1^2 + b*x1*x2 ......

I don't want to have terms with second degree like x1^2.

how can I get

y = a*x1 + b*x2 + c*x3 + d*x1*x2

if x1 and x2 have high correlation larger than some threshold value j .

221

asked Aug 23 '17 00:08

Dylan

2 Answers

For generating polynomial features, I assume you are using sklearn.preprocessing.PolynomialFeatures

There's an argument in the method for considering only the interactions. So, you can write something like:

poly = PolynomialFeatures(interaction_only=True,include_bias = False) poly.fit_transform(X)

Now only your interaction terms are considered and higher degrees are omitted. Your new feature space becomes [x1,x2,x3,x1*x2,x1*x3,x2*x3]

You can fit your regression model on top of that

clf = linear_model.LinearRegression() clf.fit(X, y)

Making your resultant equation y = a*x1 + b*x2 + c*x3 + d*x1*x + e*x2*x3 + f*x3*x1

Note: If you have high dimensional feature space, then this would lead to curse of dimensionality which might cause problems like overfitting/high variance

116

answered Sep 28 '22 16:09

harsha

Use patsy to construct a design matrix as follows:

y, X = dmatrices('y ~ x1 + x2 + x3 + x1:x2', your_data)

Where your_data is e.g. a DataFrame with response column y and input columns x1, x2 and x3.

Then just call the fit method of your estimator, e.g. LinearRegression().fit(X,y).

answered Sep 28 '22 16:09

DontDivideByZero

Related questions
                            
                                Convert pandas series of lists to dataframe
                            
                                Lodash: is it possible to use map with async functions?
                            
                                What is the difference between xgb.train and xgb.XGBRegressor (or xgb.XGBClassifier)?
                            
                                pip install -U setuptools fail windows 10
                            
                                ModelMapper: matches multiple source property hierarchies
                            
                                Vuetify v-select onchange event returns previously selected value instead of current
                            
                                Failed to resolve: com.google.android.material:material:1.0.0-alpha1
                            
                                Angular - assign custom validator to a FormGroup
                            
                                Specify specific props and accept general HTML props in Typescript React App
                            
                                Navigation graph with multiple top level destinations
                            
                                Javascript: Merge Two Arrays of Objects, Only If Not Duplicate (Based on Specified Object Key)
                            
                                GooglePlay - wrong signing key for app bundle [duplicate]

Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!

Donate Us With