Summary

Data Science Fundamentals Samenvatting

Rating

Sold

Pages

Uploaded on

01-06-2025

Written in

2024/2025

Summary of 22 pages for the course Data science at tmhs

Institution

Course

Whoops! We can’t load your doc right now. Try again or contact support.

Report Copyright Violation

Written for

Institution: Thomas More Hogeschool (tmhs)
Study: Applied Data Intelligence
Course: Data science (YP0861)

All documents for this subject (1)

Document information

Uploaded on: June 1, 2025
Number of pages: 22
Written in: 2024/2025
Type: Summary

Subjects

data science fundamentals
machine learning
clustering
linear regression
polynomial regression
logistic regression
decision tree
knn
kmeans
hierarchical clustering
dimensionality reduction
lasso regres

Content preview

Summary: Data Science
Intro
1. OSEMN process

2. Machine Learning
 An approach to achieve artificial intelligence through systems that can learn from
experience to find patterns in a set of data.
 It relies on teaching a computer to recognize patterns by example, rather than
programming it with specific rules
 A way to make predictions
o Takes in data
o Learns patterns from said data
o Classifies new data it has not seen before

2.1 Types of ML
 Supervised
o Training data is labeled
o System knows expected output label
 Unsupervised
o Training data is unlabeled
o We don’t know the output

2.2 Methods of ML

2.2.1 Regression
 The variable we wish to predict (the dependent variable) is of a continuous nature.
 The value of a given entry is determined based on known cases
 Supervised

,2.2.2 Classification
 The variable we wish to predict is of a categorical nature.
 The label of a given entry is determined based on known labelled cases.
 Supervised

2.2.3 Clustering
 Detect clusters of observations in our dataset.
 A given entry is assigned to a group base on the entire dataset
 Unsupervised

3. Notebook
 n_obs
o Number of observations or data
points that you want to generate.
 x = np.linspace(-3, 3, n_obs)
o Generates n_obs points evenly spaced
between -3 and 3.
 X = x[:, np.newaxis]
o Reshapes x into a 2D array (for
compatibility in some models or
algorithms).
 y = x + x * np.random.normal(2, 0.5, n_obs):
o Generates the y values by adding random noise to x.
o Noise comes from a normal distribution with a mean of 2 and a standard
deviation of 0.5.
o Adds variability to the relationship between x and y

 Make function (Linear Regression) object to implement this algorithm
o regressor = LinearRegression()
 Run the OLS algorithm in order to fit the function on our data
o regressor.fit(X, y)

, Linear Regression
1. Regression
 In a regression problem we try to understand the behavior (read: analyse / predict) of a
certain (continuous) variable (dependent variable) by studying the influence another
variable (independent variable) has on it.
 We want to predict Y based on X
o Does “hours studied” affect the variable “exam grade”?
o Does “age” affect “income”?
o Does “muscle mass” affect “time to run a marathon”?
o Does “advertising budget” affect “products sold”?

2. Linear Regression
 The simplest form of regression
 A linear model → a straight line through the data
 The higher X, the higher (or lower) Y
 “line of best fit”

3. Linear relation = linear function
 Mathematical function
o 𝑓(𝑥) = 𝑎𝑥 + 𝑏
o 𝑦𝑖 = 𝛽 0+ 𝛽 1𝑥
 Beta 0 is the intercept
 Where the function crosses the X-axis • Value of
Y when X = 0
 Beta 1 is the slope
 Postive Beta 1 → the function grows
 Negative Beta 1 → the function lowers
 The increase amount Y with each increase of X

4. Multiple Linear Regression
 Same as linear regression, but with multiple factors
o Ex: “Income” is affected by “seniority” and “years of education”
 “Plane of best fit”
 What happens with our function?
 Our intercept remains
 A new “slope” is created for each parameter
o 𝑓(𝑥) = 𝑎𝑥 + 𝑏𝑥 + 𝑐𝑥 + 𝑑𝑥 + … + 𝑒
o 𝑦𝑖 = 𝛽0 + 𝛽1𝑥 + 𝛽2𝑥 + 𝛽3𝑥 + 𝛽4𝑥 + …

5. Model training
 We split our data

5.1 Train – Test split

$9.58

Get access to the full document:

100% satisfaction guarantee

Immediately available after payment

Both online and in PDF

No strings attached

Get to know the seller

soetenssimon

Get to know the seller

soetenssimon Thomas More Hogeschool

View profile

Sold

Member since

10 months

Number of followers

Documents

Last sold

0.0

0 reviews

Why students choose Stuvia

Created by fellow students, verified by reviews

Quality you can trust: written by students who passed their tests and reviewed by others who've used these notes.

Didn't get what you expected? Choose another document

No worries! You can instantly pick a different document that better fits what you're looking for.

Pay as you like, start learning right away

No subscription, no commitments. Pay the way you're used to via credit card and download your PDF document instantly.

“Bought, downloaded, and aced it. It really can be that simple.”

Alisha Student

Frequently asked questions

What do I get when I buy this document?

You get a PDF, available immediately after your purchase. The purchased document is accessible anytime, anywhere and indefinitely through your profile.

Satisfaction guarantee: how does it work?

Our satisfaction guarantee ensures that you always find a study document that suits you well. You fill out a form, and our customer service team takes care of the rest.

Who am I buying these notes from?

Stuvia is a marketplace, so you are not buying this document from us, but from seller soetenssimon. Stuvia facilitates payment to the seller.

Will I be stuck with a subscription?

No, you only buy these notes for $9.58. You're not tied to anything after your purchase.

Can Stuvia be trusted?

4.6 stars on Google & Trustpilot (+1000 reviews) 50201 documents were sold in the last 30 days Founded in 2010, the go-to place to buy study notes for 15 years now

Data Science Fundamentals Samenvatting

Written for

Document information

Subjects

Content preview

Get to know the seller

Recently viewed by you

Why students choose Stuvia

Created by fellow students, verified by reviews

Didn't get what you expected? Choose another document

Pay as you like, start learning right away

Frequently asked questions

What do I get when I buy this document?

Satisfaction guarantee: how does it work?

Who am I buying these notes from?

Will I be stuck with a subscription?

Can Stuvia be trusted?