I am experimenting with the Spark Vertica connector and trying to load from a Kerberos enabled Hadoop cluster to a non Kerberos Vertica cluster. I am unable to setup the HDFS scheme as the Vertica cluster doesnt support Kerberos. I am currently experimenting loading using the Vertica copy stream option using JDBC driver but would assume that the connector will have better throughput. My questions below:
1. Can the Spark Vertica connector be used to send data from a Kerberos enabled hadoop cluster to a non Kerberos Vertica cluster?
2. What is the fastest method to load data into Vertica from an external system ? (Spark-Vertica connector, Vertica Copy Stream JDBC, JDBC bulk loading, Streaming from Kafka to Vertica)
Appreciate all the input.