CCA175試験は出題範囲が広く、独学だけでは不安を感じる方も少なくありません。ShikenPASSのCloudera CCA Spark and Hadoop Developer問題集は96問の練習問題で出題傾向を網羅しており、弱点分野の克服をしっかりサポートします。
Cloudera CCA175 試験概要:
| 認定ベンダー: | Cloudera |
|---|---|
| 試験名: | CCA SparkおよびHadoop開発者試験 |
| 試験番号: | CCA175 |
| 試験形式: | 実技重視形式, 実践的課題の実施, ライブクラスタ環境での実行 |
| 出題数: | 8~12問 |
| 認定の有効期間: | 2年間 |
| 合格点: | 70% |
| 関連資格: | CCA Data Analyst CCP Data Engineer |
| 試験時間: | 120 分 |
| 受験料: | 295米ドル |
| 対応言語: | 英語 |
| 推奨トレーニング: | Cloudera SparkおよびHadoop向け開発者トレーニング |
| 受験申し込み: | Cloudera認定試験ポータル |
| サンプル問題: | DOWNLOAD DEMO |
| 受験方法: | 監督者によるオンライン監視形式で、専用のセキュアブラウザを通じてリモートで受験します |
| 前提条件: | 公式な受験資格は設けられていません。ScalaまたはPython、HDFS、Hive、Sqoop、Flume、Sparkに関する実務経験が推奨されます |
| 公式シラバスのURL: | https://www.cloudera.com/about/training/certification/cdhhdp-certification/cca-spark.html |
Cloudera CCA175 試験シラバストピック:
| セクション | 比重 | 目標 |
|---|---|---|
| データの変換、一時保管、保存 | 35% | - HDFSからSparkにデータを読み込む - Spark APIを使用してETL処理を実行する - 変換後の結果をHDFSに書き戻す - 各種ファイル形式(CSV、JSON、Avro、Parquetなど)の読み書きを行う |
| データ分析 | 35% | - 一時ビューおよびテーブルの作成とクエリを実行する - 集計処理と計算を実施する - 複数のデータセットを結合する - データセットのフィルタリングと並べ替えを行う - Spark SQLを使用し、Hiveメタストアと連携する |
| データの取り込み | 20% | - リレーショナルデータベースとHDFS間でデータのインポート・エクスポートを行う - Flumeを使用してHDFSにデータを取り込む - Sqoopを使用してHDFSにデータを取り込む |
| アプリケーションの設定と実行 | 10% | - Spark Shell(Scala/PySpark)を使用して対話的な分析を行う - spark-submitを使用してSparkアプリケーションを送信する - メモリ、コア数、実行に関する各種パラメータを設定する |
CCA175試験に関するQ&Aまとめ
CCA175試験は、Clouderaが実施する公式の認定試験で、合格すると「Cloudera Certified Associate(CCA)SparkおよびHadoop開発者」の認定を取得できます。この認定はアソシエイトレベルに位置づけられています。関連する認定資格にはCCP Data Engineer、CCA Data Analystなどがあります。詳しい試験情報は、このページの試験概要やほかの質問項目でもご紹介しています。
CCA175試験の問題数は8~12問、制限時間は120 分です。限られた時間で全問を解き切るには、1問ごとのペースを意識し、難しい問題に時間を使いすぎない進め方が重要です。ShikenPASSのテストエンジンには時間制限付きの模擬試験モードがあるため、本番と同じ時間配分で96問の練習問題に取り組めます。受験直前には、必ず時間を計った通し演習で時間感覚を確認しておきましょう。
CCA175試験の合格に必要なスコアは70%、受験料は295米ドルです。万が一不合格になった場合、再受験には受験料を再度全額支払う必要があるため、本番前の仕上がり確認が欠かせません。ShikenPASSの練習問題で模擬試験を繰り返し、安定して合格点を上回る状態になってから受験することをおすすめします。
はい。ShikenPASSではCloudera CCA Spark and Hadoop Developer(CCA175)問題集の無料サンプルをご用意しており、購入前に内容や品質をご確認いただけます。また、ご購入後は365日間無料で最新版にアップデートでき、更新期間終了後の更新も50%割引でご利用いただけます。
ShikenPASSでは「返金保証」をご用意しています。ご購入後60日以内にCCA175試験を受験して不合格だった場合、受験票の写しと公式スコアレポート(Score Report)のPDFを試験後2日以内にご提出いただければ、7日以内に全額返金の手続きが完了します。なお、ご購入後3日以内の受験や、実際に受験されなかった場合、無料資料・期限切れのご注文は対象外となり、受験者名とお支払い者名の一致が必要です。返金の代わりに、同等の試験資料2点を無料でお受け取りいただき、元の製品の更新サービスを継続することも可能です。商品はお支払い完了後すぐにダウンロードでき、1分以内にメールでもお届けします。2時間経っても届かない場合はカスタマーサポートまでご連絡ください。インストールできるパソコンの台数に制限はありません。
CCA175試験の出題範囲は、公式の発表では全部で4の分野に分かれています。主な分野としては、「アプリケーションの設定と実行」(10%)、「データの取り込み」(20%)、「データの変換、一時保管、保存」(35%)などが挙げられます。各分野に含まれる詳細なトピックは、上記の試験範囲一覧でご確認ください。
Cloudera CCA Spark and Hadoop Developer 認定 CCA175 試験問題:
問題 #1
CORRECT TEXT
Problem Scenario 65 : You have been given below code snippet.
val a = sc.parallelize(List("dog", "cat", "owl", "gnu", "ant"), 2)
val b = sc.parallelize(1 to a.count.tolnt, 2)
val c = a.zip(b)
operation1
Write a correct code snippet for operationl which will produce desired output, shown below.
Array[(String, Int)] = Array((owl,3), (gnu,4), (dog,1), (cat,2>, (ant,5))
問題 #2
CORRECT TEXT
Problem Scenario 82 : You have been given table in Hive with following structure (Which you have created in previous exercise).
productid int code string name string quantity int price float
Using SparkSQL accomplish following activities.
1 . Select all the products name and quantity having quantity <= 2000
2 . Select name and price of the product having code as 'PEN'
3 . Select all the products, which name starts with PENCIL
4 . Select all products which "name" begins with 'P\ followed by any two characters, followed by space, followed by zero or more characters
問題 #3
CORRECT TEXT
Problem Scenario 96 : Your spark application required extra Java options as below. -
XX:+PrintGCDetails-XX:+PrintGCTimeStamps
Please replace the XXX values correctly
./bin/spark-submit --name "My app" --master local[4] --conf spark.eventLog.enabled=talse -
-conf XXX hadoopexam.jar
問題 #4
CORRECT TEXT
Problem Scenario 32 : You have given three files as below.
spark3/sparkdir1/file1.txt
spark3/sparkd ir2ffile2.txt
spark3/sparkd ir3Zfile3.txt
Each file contain some text.
spark3/sparkdir1/file1.txt
Apache Hadoop is an open-source software framework written in Java for distributed storage and distributed processing of very large data sets on computer clusters built from commodity hardware. All the modules in Hadoop are designed with a fundamental assumption that hardware failures are common and should be automatically handled by the framework spark3/sparkdir2/file2.txt
The core of Apache Hadoop consists of a storage part known as Hadoop Distributed File
System (HDFS) and a processing part called MapReduce. Hadoop splits files into large blocks and distributes them across nodes in a cluster. To process data, Hadoop transfers packaged code for nodes to process in parallel based on the data that needs to be processed.
spark3/sparkdir3/file3.txt
his approach takes advantage of data locality nodes manipulating the data they have access to to allow the dataset to be processed faster and more efficiently than it would be in a more conventional supercomputer architecture that relies on a parallel file system where computation and data are distributed via high-speed networking
Now write a Spark code in scala which will load all these three files from hdfs and do the word count by filtering following words. And result should be sorted by word count in reverse order.
Filter words ("a","the","an", "as", "a","with","this","these","is","are","in", "for",
"to","and","The","of")
Also please make sure you load all three files as a Single RDD (All three files must be loaded using single API call).
You have also been given following codec
import org.apache.hadoop.io.compress.GzipCodec
Please use above codec to compress file, while saving in hdfs.
問題 #5
CORRECT TEXT
Problem Scenario 7 : You have been given following mysql database details as well as other info.
user=retail_dba
password=cloudera
database=retail_db
jdbc URL = jdbc:mysql://quickstart:3306/retail_db
Please accomplish following.
1. Import department tables using your custom boundary query, which import departments between 1 to 25.
2 . Also make sure each tables file is partitioned in 2 files e.g. part-00000, part-00002
3 . Also make sure you have imported only two columns from table, which are department_id,department_name
解説:
| 問題 #1 正解: 会員のみ閲覧可能 | 問題 #2 正解: 会員のみ閲覧可能 | 問題 #3 正解: 会員のみ閲覧可能 | 問題 #4 正解: 会員のみ閲覧可能 | 問題 #5 正解: 会員のみ閲覧可能 |

弊社は製品に自信を持っており、面倒な製品を提供していません。


-冈田**

