CCA175 CCA Spark and Hadoop Developer Exam
Structured Study Notes and Topic Review
CCA175 CCA SPARK AND HADOOP DEVELOPER EXAM
STRUCTURED STUDY NOTES & TOPIC REVIEW
Based on Official Cloudera Certification Guide
SECTION 1: EXAM OVERVIEW
The CCA175 is a practical, hands-on exam where candidates complete tasks using
a live environment. You have 120 minutes to complete the exam [citation:8].
Candidates should be familiar with:
- Importing/exporting data using Sqoop
- Loading and transforming data using Spark (Scala or Python)
- Writing Spark SQL queries
- Working with Hive metastore tables
- Configuring Spark applications [citation:3][citation:11]
1
,SECTION 2: DATA INGESTION (SQOOP & FLUME)
QUESTION 1
Which Sqoop argument is used to import a table as Avro data files?
A) --as-avrodatafile
B) --as-parquetfile
C) --as-textfile
D) --as-sequencefile
Correct Answer: A
Rationale: The --as-avrodatafile argument directs Sqoop to import data in Avro
format. This is a specific file format option used during Sqoop imports
[citation:2][citation:10].
QUESTION 2
2
,You need to import data from a MySQL database into HDFS and change the
delimiter from comma to tab during import. Which Sqoop argument should you
use?
A) --fields-terminated-by '\t'
B) --delimiter '\t'
C) --input-fields-terminated-by '\t'
D) --target-dir
Correct Answer: A
Rationale: The --fields-terminated-by argument allows you to specify a custom
field delimiter for the imported data. Sqoop also supports changing file formats
during import, such as Avro or Parquet [citation:3][citation:11][citation:10].
QUESTION 3
You have processed data in HDFS and need to export the results back to a MySQL
table. Which Sqoop command and arguments would you use?
A) sqoop import --connect jdbc:mysql://localhost/retail_db --username retail_dba -
-password cloudera --table result --export-dir /user/cloudera/problem1/result4a-csv
B) sqoop export --connect jdbc:mysql://localhost/retail_db --username retail_dba --
password cloudera --export-dir /user/cloudera/problem1/result4a-csv --table result
3
, C) sqoop export --connect jdbc:mysql://localhost/retail_db --username retail_dba --
password cloudera --table result --import-dir /user/cloudera/problem1/result4a-csv
D) sqoop import --connect jdbc:mysql://localhost/retail_db --username retail_dba -
-password cloudera --table result --input-dir /user/cloudera/problem1/result4a-csv
Correct Answer: B
Rationale: The `sqoop export` command is used to export from HDFS to a
relational database. The `--export-dir` parameter specifies the HDFS directory
containing the data to export, and `--table` specifies the target table in the database
[citation:10].
QUESTION 4
You are working with a Sqoop import. The database connection parameters are
missing. Which of the following is the standard JDBC connection string format for
Sqoop?
A) jdbc:mysql://hostname:port/database
B) mysql://hostname:port/database
C) sqoop:mysql://hostname:port/database
D) jdbc:mysql://hostname/database:port
Correct Answer: A
4
Structured Study Notes and Topic Review
CCA175 CCA SPARK AND HADOOP DEVELOPER EXAM
STRUCTURED STUDY NOTES & TOPIC REVIEW
Based on Official Cloudera Certification Guide
SECTION 1: EXAM OVERVIEW
The CCA175 is a practical, hands-on exam where candidates complete tasks using
a live environment. You have 120 minutes to complete the exam [citation:8].
Candidates should be familiar with:
- Importing/exporting data using Sqoop
- Loading and transforming data using Spark (Scala or Python)
- Writing Spark SQL queries
- Working with Hive metastore tables
- Configuring Spark applications [citation:3][citation:11]
1
,SECTION 2: DATA INGESTION (SQOOP & FLUME)
QUESTION 1
Which Sqoop argument is used to import a table as Avro data files?
A) --as-avrodatafile
B) --as-parquetfile
C) --as-textfile
D) --as-sequencefile
Correct Answer: A
Rationale: The --as-avrodatafile argument directs Sqoop to import data in Avro
format. This is a specific file format option used during Sqoop imports
[citation:2][citation:10].
QUESTION 2
2
,You need to import data from a MySQL database into HDFS and change the
delimiter from comma to tab during import. Which Sqoop argument should you
use?
A) --fields-terminated-by '\t'
B) --delimiter '\t'
C) --input-fields-terminated-by '\t'
D) --target-dir
Correct Answer: A
Rationale: The --fields-terminated-by argument allows you to specify a custom
field delimiter for the imported data. Sqoop also supports changing file formats
during import, such as Avro or Parquet [citation:3][citation:11][citation:10].
QUESTION 3
You have processed data in HDFS and need to export the results back to a MySQL
table. Which Sqoop command and arguments would you use?
A) sqoop import --connect jdbc:mysql://localhost/retail_db --username retail_dba -
-password cloudera --table result --export-dir /user/cloudera/problem1/result4a-csv
B) sqoop export --connect jdbc:mysql://localhost/retail_db --username retail_dba --
password cloudera --export-dir /user/cloudera/problem1/result4a-csv --table result
3
, C) sqoop export --connect jdbc:mysql://localhost/retail_db --username retail_dba --
password cloudera --table result --import-dir /user/cloudera/problem1/result4a-csv
D) sqoop import --connect jdbc:mysql://localhost/retail_db --username retail_dba -
-password cloudera --table result --input-dir /user/cloudera/problem1/result4a-csv
Correct Answer: B
Rationale: The `sqoop export` command is used to export from HDFS to a
relational database. The `--export-dir` parameter specifies the HDFS directory
containing the data to export, and `--table` specifies the target table in the database
[citation:10].
QUESTION 4
You are working with a Sqoop import. The database connection parameters are
missing. Which of the following is the standard JDBC connection string format for
Sqoop?
A) jdbc:mysql://hostname:port/database
B) mysql://hostname:port/database
C) sqoop:mysql://hostname:port/database
D) jdbc:mysql://hostname/database:port
Correct Answer: A
4