Hi Chandra,
To get you started, for large sets of documents you will likely want to make use of the FileNet bulk sweep jobs to export the relevant documents from the repository utilizing FileNet background cycles. There is a Knowledgecenter topic for Datacap 9.1.7 related to how a FileNet bulk sweep can be configured for export and consumption by Datacap:
https://www.ibm.com/support/knowledgecenter/SSZRWV_9.1.7/com.ibm.dc.develop.doc/dcdev629.htmOnce you have processed the documents in Datacap you will want to make use of the FNP8_UpdateContent action (I believe the example in the above link shows how to update the metadata properties but not version the content itself). Here is a link to the relevant documentation:
https://www.ibm.com/support/knowledgecenter/SSZRWV_9.1.7/com.ibm.dc.reference.doc/dcaca938.htmRegarding hardware configuration, if you have available cycles the bulk sweep with utilize background cycles and can be scheduled in off-peak hours so as to not affect FileNet responsiveness. You will have to take into account the duration over which you will OCR said documents with Datacap (as OCR is one of the most CPU intensive operations) and how much new ingestion headroom you have for FileNet as the upload will occur over the FileNet WSI web service.
I hope this helps you get started.
Thanks,
Tim
------------------------------
TIM PASCARELLA
------------------------------