100% satisfaction guarantee Immediately available after payment Both online and in PDF No strings attached 4.2 TrustPilot
logo-home
Summary

JADS Master - Data Engineering Summary

Rating
-
Sold
4
Pages
24
Uploaded on
02-01-2022
Written in
2020/2021

Summary for the Data Engineering course of the Master Data Science and Entrepreneurship.

Institution
Course









Whoops! We can’t load your doc right now. Try again or contact support.

Written for

Institution
Study
Course

Document information

Uploaded on
January 2, 2022
Number of pages
24
Written in
2020/2021
Type
Summary

Subjects

Content preview

1. Introduction
Data Engineer
Develops the architecture that helps analyse and process data in the way the organization
needs it.




Data Science Lifecycle




Big Data
Term for a collection of datasets so large and complex that it becomes difficult to process using
traditional data processing applications.

Structured Data Semi-Structured Data Unstructured Data
RDMS XML, RDF, JSON, etc. Video, Images, Text, etc.

V3 Model
● Volume: enterprises are always growing in terms of data
● Velocity: sometimes 2 minutes is too late, data must be used as streams in order to
maximize its value
● Variety: structured and unstructured data, new insights are found when analyzing these
data types together

Data Pipeline
Aggregates, organizes and moves data for storage, insights and analysis.




1

, ML Ops
● Software engineering approach: enables team to efficiently produce high quality software
● Cross-functional team: experts with different skill sets and workflows (DE, DS, ML, Dev,
Ops)
● Producing software based on code, data and models: all artifacts of ML software
production process require different tools and workflows → must be versioned and
managed accordingly
● Small and safe increments: the release of software artifacts is divided into small
increments → allows visibility and control around the levels of variance of its outcomes
● Reproducible and reliable software release: model outputs are non-deterministic →
process of releasing ML software is reliable and reproducible → leverage automation as
much as possible
● Software release at any time: ML software needs to be delivered into production at any
time → when to release is a business decision rather than a technical decision
● Short adaptation cycles: short cycles mean development cycles are in the order of days
or hours




2. Cloud Computing & Virtualization
Cloud
A type of distributed system consisting of interconnected and virtualized computers dynamically
provisioned and presented as one (or more) unified computing resource(s) based on
service-level agreements established through negotiation between service providers and
consumers.
● Cloud contains “your” data
● Cloud computes “your” data
● Cloud hosts “your” data intensive applications

▶ Central ideas:
● Utility computing over data
● SOA (Service Oriented Architectures)
● SLA (Service Level Agreements)



2

Get to know the seller

Seller avatar
Reputation scores are based on the amount of documents a seller has sold for a fee and the reviews they have received for those documents. There are three levels: Bronze, Silver and Gold. The better the reputation, the more your can rely on the quality of the sellers work.
tomdewildt Jheronimus Academy of Data Science
Follow You need to be logged in order to follow users or courses
Sold
29
Member since
4 year
Number of followers
13
Documents
22
Last sold
6 months ago

5.0

1 reviews

5
1
4
0
3
0
2
0
1
0

Recently viewed by you

Why students choose Stuvia

Created by fellow students, verified by reviews

Quality you can trust: written by students who passed their tests and reviewed by others who've used these notes.

Didn't get what you expected? Choose another document

No worries! You can instantly pick a different document that better fits what you're looking for.

Pay as you like, start learning right away

No subscription, no commitments. Pay the way you're used to via credit card and download your PDF document instantly.

Student with book image

“Bought, downloaded, and aced it. It really can be that simple.”

Alisha Student

Frequently asked questions