KnowledgeHub
Questions
Tags
Users
Search
Alex Rivera
|
Logout
Edit Question
Title
Body
The file that I am loading is separated by ' ' (white space). Below is the file. The file resides in HDFS:- 001 000 001 000 002 001 003 002 004 003 005 004 006 005 007 006 008 007 099 007 1> I am creating an external table and loading the file by issuing the below command:- CREATE EXTERNAL TABLE IF NOT EXISTS graph_edges (src_node_id STRING COMMENT 'Node ID of Source node', dest_node_id STRING COMMENT 'Node ID of Destination node') ROW FORMAT DELIMITED FIELDS TERMINATED BY ' ' STORED AS TEXTFILE LOCATION '/user/hadoop/input'; 2> After this, I am simply inserting the table in another file by issuing the below command:- INSERT OVERWRITE DIRECTORY '/user/hadoop/output' SELECT * FROM graph_edges; 3> Now, when I cat the file, the fields are not separated by any delimiter:- hadoop dfs -cat /user/hadoop/output/000000_0 Output:- 001000 001000 002001 003002 004003 005004 006005 007006 008007 099007 Can somebody please help me out? Why is the delimiter being removed and how to delimit the output file? In the CREATE TABLE command I tried DELIMITED BY '\t' but then I am getting unnecessary NULL column. Any pointers help much appreciated. I am using Hive 0.9.0 version.
Tags (comma-separated)
Save Edits
Cancel