Class notes

Stanford CS229 Notes - Regression Algorithms

Rating

Sold

Pages

Uploaded on

02-01-2025

Written in

2024/2025

1. Introduction to Linear Regression and Gradient Descent Purpose: Introduces linear regression as a foundational supervised learning algorithm. Content Highlights: Explanation of hypothesis formulation. Detailed notation and definitions (parameters, input vectors, target variables). Step-by-step derivation of cost function

Show more Read less

Institution

Course

Content preview

Stanford CS229: Machine Learning
Amrit Kandasamy
November 2024

1 Linear Regression and Gradient Descent
Lecture Note Slides

1.1 Notation and Definitions
Pn
Linear Regression Hypothesis: hθ (x) = i=0 θi xi , where x0 = 1.
 
θ0
 .. 
θ= . 
θn
is called the parameters of the learning algorithm. The algorithm’s job is to
choose θ.
 
x0
 .. 
x= . 
xn
is an input vector (often the inputs are called features).
We let m be the number of training examples (elements in the training set).
y is the output, sometimes called the target variable.
(x, y) is one training example. We will use the notation

(x(i) , y (i) )

to denote the ith training example.
As used in the vectors and summation n is the number of features.
:= denotes assignment (usually of some variable or function). For example,
a := a + 1 increments a by 1.
We write hθ (x) as h(x) for convenience.

1

, Figure 1: Visual of Gradient Descent with Two Parameters

1.2 How to Choose Parameters θ
Choose θ such that h(x) ≈ y for the training examples. Generally, we want to
minimize
m
1X
J(θ) = (hθ (x(i) ) − y)2
2 i=1

In order to minimize J(θ), we will employ Batch Gradient Descent.
Let’s look an example with 2 parameters. Start with some point (θ0 , θ1 , J(θ)),
determined either randomly or by some condition. We look around all around
and think,

”What direction should we take a tiny step in to go downward as fast as possible?”.

If a different starting point was used, the resulting optimum minima would have
been changed (see the two paths above).

Now let’s formalize the gradient descent algorithm(s).

1.2.1 Batch Gradient Descent
Let α be the learning rate. Then the algorithm can be written as

∂
θj := θj − α J(θ)
∂θj

Let’s derive the partial derivative part. Assume there’s only 1 training example
for now. Substituting our definition of J, we have
n
!
∂ ∂ 1 2 ∂ X
α J(θ) = (hθ (x) − y) = (hθ (x) − y) · ( θ i xi ) − y
∂θj ∂θj 2 ∂θj i=0

2

Report Copyright Violation

Written for

Institution: Stanford University
Course: CS229

All documents for this subject (7)

Document information

Uploaded on: January 2, 2025
File latest updated on: January 2, 2025
Number of pages: 12
Written in: 2024/2025
Type: Class notes
Professor(s): Unknown
Contains: All classes

Subjects

algorithm
machine learning
regression
classification
artificial intelligence
cheap
linear model
high quality
examples
cs229
stanford
glas
perceptron
biology
physics
science
notes

R269,02

Get access to the full document:

Written by students who passed

Immediately available after payment

Read online or as PDF

Get to know the seller

tuningnumbers

Get to know the seller

tuningnumbers stanford university

View profile

Sold

Member since

1 year

Number of followers

Documents

Last sold

0,0

0 reviews

Why students choose Stuvia

Created by fellow students, verified by reviews

Quality you can trust: written by students who passed their exams and reviewed by others who've used these notes.

Didn't get what you expected? Choose another document

No worries! You can immediately select a different document that better matches what you need.

Pay how you prefer, start learning right away

No subscription, no commitments. Pay the way you're used to via credit card or EFT and download your PDF document instantly.

“Bought, downloaded, and aced it. It really can be that simple.”

Alisha Student

Frequently asked questions

What do I get when I buy this document?

You get a PDF, available immediately after your purchase. The purchased document is accessible anytime, anywhere and indefinitely through your profile.

Satisfaction guarantee: how does it work?

Our satisfaction guarantee ensures that you always find a study document that suits you well. You fill out a form, and our customer service team takes care of the rest.

Who am I buying this summary from?

Stuvia is a marketplace, so you are not buying this document from us, but from seller tuningnumbers. Stuvia facilitates payment to the seller.

Will I be stuck with a subscription?

No, you only buy this summary for R269,02. You're not tied to anything after your purchase.

Can Stuvia be trusted?

4.6 stars on Google & Trustpilot (+1000 reviews) 50161 documents were sold in the last 30 days Founded in 2010, the go-to place to buy summaries for 16 years now

Stanford CS229 Notes - Regression Algorithms

Content preview

Written for

Document information

Subjects

More courses for Stanford University >

Get to know the seller

Trending documents

Why students choose Stuvia

Created by fellow students, verified by reviews

Didn't get what you expected? Choose another document

Pay how you prefer, start learning right away

Frequently asked questions

What do I get when I buy this document?

Satisfaction guarantee: how does it work?

Who am I buying this summary from?

Will I be stuck with a subscription?

Can Stuvia be trusted?