#757 introducing article type extraction along with unit test. Article type will be required for filtering out pmc duplicates and leaving only proper types
introducing cloudera repository in parent container, removing repository definitions from individual IIS modules
fixing sourceDocumentId which is now extracted from input DocumentText record conveying NLM
#757 fixing pmc citation matching test by providing proper input
#757 fixing pmid and doi matching, fixing sourceDocumentId and destinationDocumentId generation
Commented out test in a stub of a solution to the task #576: Ingestion of metadata from EuropePMC.
Stub of a solution to the task #576: Ingestion of metadata from EuropePMC.
Refactored code to use the XPathEvaluator.fromString method.
updating default job properties
renaming workflow to ingest_pmc_plaintext
Excluding conflicting dependency
replacing "result" string with Type.result.name()
updating job.properties
dir names in parameters should not contain nameNode
rename a field
introducing deploy.info file for module icm-iis-ingest-pmc