Fusioninsight HDofhiveIn the application, there are the following scenarios: high compression efficiency for data storage files rate, and most queries involve only a subset of fields in the file. This scenario is suitable for using columnar files (ORC File) storage
existSparkmiddle,SparksQLis an independent module that does not depend onSparkCorefinish independentlySQL Actions such as the corner line of the statement.
A large-scale production enterprise,tPlan to analyze the internal logistics data and sales datat,Assumeiout of the wayT effectganalysis, which of the following statements are correct? (multiple choice)
existMapReduceDuring application development,setMapOutputCompressorClassWhat is the role of classes?
FusionInsight HDin, yesSolrThe creation of various resources and the use of read and write permissions, which of the following statements is wrong?
pass throughHBasefofcreateTableThe method creates a table, what parameters must be passed in?
aboutKafkaofProducer, is the following statement correct? (multiple choice)
existMapReduceIn the development framework,InputFormatWhat is the function of the class?
HDFSIt adopts a "write once, read many" file access model. So it is recommended that a file be created, written and After closing, do not modify it again.
FusionInsight HDin, aboutHivethe data load function (viaHiveofLOADcommand guide input data), which of the following descriptions is wrong?
Sparkis a memory-based computing engine, allSparkData during program operation can only be stored in in memory.
deployFusionInsight HD within the same clusterFlume ServerHow many nodes are recommended to deploy at least?