Configuring the connection to the file system to be used by Spark
-
Double-click tHDFSConfiguration to open its
Component view. -
In the Version area, select the Hadoop
distribution you need to connect to and its version. -
In the NameNode URI
field, enter the location of the machine hosting the NameNode service of the
cluster. If you are using WebHDFS, the location should be
webhdfs://masternode:portnumber; if this WebHDFS is secured
with SSL, the scheme should be swebhdfs and you need to use
a tLibraryLoad in the Job to load the library required by
the secured WebHDFS. -
In the Username field, enter the authentication
information used to connect to the HDFS system to be used.
Document get from Talend https://help.talend.com
Thank you for watching.
Subscribe
Login
0 Comments