Question 11

Which of the following is NOT one of the three main types of triggers that Dataflow supports?
  • Question 12

    You are building a new application that you need to collect data from in a scalable way. Data arrives continuously from the application throughout the day, and you expect to generate approximately 150 GB of JSON data per day by the end of the year. Your requirements are:
    Decoupling producer from consumer
    Space and cost-efficient storage of the raw ingested data, which is to be stored indefinitely
    Near real-time SQL query
    Maintain at least 2 years of historical data, which will be queried with SQ
    Which pipeline should you use to meet these requirements?
  • Question 13

    Which of the following is NOT one of the three main types of triggers that Dataflow supports?
  • Question 14

    As your organization expands its usage of GCP, many teams have started to create their own projects.
    Projects are further multiplied to accommodate different stages of deployments and target audiences. Each project requires unique access control configurations. The central IT team needs to have access to all projects.
    Furthermore, data from Cloud Storage buckets and BigQuery datasets must be shared for use in other projects in an ad hoc way. You want to simplify access control management by minimizing the number of policies.
    Which two steps should you take? (Choose two.)
  • Question 15

    You want to automate execution of a multi-step data pipeline running on Google Cloud. The pipeline includes Cloud Dataproc and Cloud Dataflow jobs that have multiple dependencies on each other. You want to use managed services where possible, and the pipeline will run every day. Which tool should you use?