Colleagues,
Apologies if this sounds a rather dumb question, but if one can use the generic Spark JDBC connector to write directly to Vertica (accepting its limitations), I was wondering why do we have to “…have an HDFS cluster for an intermediate staging location [to save data from Spark to Vertica] when using our Vertica Connector for Apache Spark/JDBC client library?
Is it not possible to stage the data outside HDFS (when using our Connector / JDBC)?
Regards
Mark