Is there a light-weight solution to change the datatype of specific column in ORC file without having to convert entire column datatype and re-writing entire orc file?
The following is a heavy-weight solution:
- Read orc file in Spark
- Convert datatype of a specific column
- Write converted orc file to HDFS
Looking for a light-weight solution where I can just alter embedded metadata info.
Thanks!