ODI_Solutions_For_BigData_On_Hadoop

整理文档很辛苦,赏杯茶钱您下走!

免费阅读已结束,点击下载阅读编辑剩下 ...

阅读已结束,您可以下载文档离线阅读编辑

资源描述

ODI–HadoopLabExercisesIntroduction..................................................................................................................................................4Pre-Requisites...............................................................................................................................................4Lab-1-LoadunstructureddatafromFile(LocalfilesystemorHDFS)intoHive.....................................7UnderstandingtheLab:............................................................................................................................7Executingthelab.....................................................................................................................................13Lab2-TransformandvalidatestructureddataonHive............................................................................15UnderstandingtheLab:..........................................................................................................................16Executingthelab.....................................................................................................................................18Lab3-TransformunstructureddataonHive.............................................................................................19UnderstandingtheLab:..........................................................................................................................20Executingthelab.....................................................................................................................................21Lab4-LoaddatafromHadoop(HDFSorHivetable)toOracleusingOracleLoaderforHadoop..........22UnderstandingtheLab:..........................................................................................................................22Executingthelab.....................................................................................................................................24IntroductionTheODIsolutionforBigDatalabwork,broadlyconsistsoffourlabs,eachofwhichhighlightsthecapabilitiesofODIKnowledgeModuleswhicharedevelopedtosimplifytheprocessingofBigDataonHadoop.Itshouldbenoted,however,thelabsdonotprovideacompletecoverageonalltheoptionsavailableinODIandtheknowledgemodules.Itishighlyrecommendedtousethelabsasastartingpointandrefertothedocumentation–ODIApplicationAdaptorforHadoopforadditionaldetailstodevelopcustomsolutions.SincethefocusonthelabsessionisonODIsolutionsforBigDataonHadoop,ItisassumedthatuserhasaknowledgeofODIandHadooptechnology,specificallyHDFSandHive.TheknowledgemodulesweredevelopedtosimplifyprocessingofunstructuredandstructureddataonHadoop.TheknowledgemodulesweredevelopedtofacilitatethefollowingprocessTheknowledgemodulesthathelpinfacilitatingtheaboveprocessare:LoadunstructureddatafromFile(LocalfilesystemorHDFS)intoHive–Box1TransformandvalidatestructureddataonHive–Box2TransformunstructureddataonHive–Box2LoadprocesseddatainHivetoOracle–Box3Pre-RequisitesTheODIsolutionforBigDataonHadoopisdeliveredthroughVMwithalltherequiredsoftwareandtheenvironmentpre-configured.AfterverifyingtheVMisupandrunningdothefollowing:BringupacommandterminalandenterthecommandtostartthehiveserverhiveserverBringupanothercommandterminalandenterthecommandtostartODIodiLoadDataIntoHiveTransformandvalidateDatainHiveLoadProcesseddatafromHiveintoOracleoClickon‘ConnectToRepository’inthe‘DesignerTab’oSelectOKonthepopupboxAfterselectingOK,youwillseethefollowingprojects.Thegeneralstepstoconfiguringallofthelabsarethefollowing.Sincethelabsarepre-configuredyouwillnotbeneededtoperformthesesteps,butitisimportanttobeawareofthesestepswhendevelopingcustomizedsolutions.Thedetailsofthesestepswillbeillustratedineachofthelabs.Createthemodelmeta-datawhichwillcontainthedatastoresandtheassociationtothelogicalschema.Duringexecution,dependinguponthecontextselected,thelogicalschemawillbemappedtotheappropriatephysicalschema.Thisallowsyoutoswitchcontextfromdevelopmenttotesttodeployment.Forthepurposesofthislabexercisethecontext‘global’willbeusedthroughout.Thedatastoremeta-datadescribeeitherthesourceortargetdatastore.Thedatastoresareusedintheinterfacetospecifythesourceandtarget.Definethelogicalarchitecture,physicalarchitectureandtheexecutionContextstoassociatelogicalandphysicalarchitecture.Designtheinterface.Theinterfacespecifiesthesource,target,mappingsandtherules,theKMs.Executingtheinterfacewithaspecifiedcontextwillmigratethedatafromthesourcephysicallocationintothetargetphysicallocation.Designthepackage.Thepackageallowstocreateaprocesswhichwouldexecuteinterfaces,proceduresandotherlogicasrequired.Thepackagesdesignedforthelabincludesthestepsthatgeneratedataandvalidateresults.ThegoaloftheeachofthelabsexerciseshighlighttheKnowledgemoduleslistedabove.Forthepurposesofthislabexercise,thestepstounderstandingthelabsandexecutinglargelyremainsthesame.Thereforeeachlabhastwoparts–UnderstandingtheLabandExecutingtheLab.Itis,however,importanttounderstandthedeeperpurposeofeachoftheKMsandusageofitsoptions.Whenexecuti

1 / 25
下载文档,编辑使用

©2015-2020 m.777doc.com 三七文档.

备案号:鲁ICP备2024069028号-1 客服联系 QQ:2149211541

×
保存成功