Business
Jobs
  • About Us
  • Solutions
    • Job Postings
      Post your job and receive qualified candidates in 48h.
    • Candidate Assessments
      500+ technical and psychological tests, plus anti-fraud.
    • Headhunting
      Tailor-made executive search from start to finish.
    • Payroll + EOR
      Payroll dispersal and EOR across 15+ LATAM countries.
  • Pricing
  • Jobs

0

331
Views
¿Cómo consultar datos del archivo gz de Amazon S3 usando la consulta Qubole Hive?

Necesito obtener datos específicos de gz. como escribir el sql? ¿Puedo simplemente sql como base de datos de tabla?:

 Select * from gz_File_Name where key = 'keyname' limit 10.

pero siempre regresa con un error.

over 4 years ago · Santiago Trujillo
1 answers
Answer question

0

Debe crear una tabla externa de Hive sobre esta ubicación de archivo (carpeta) para poder consultar usando Hive. Hive reconocerá el formato gzip. Como esto:

 create external table hive_schema.your_table ( col_one string, col_two string ) stored as textfile --specify your file type, or use serde LOCATION 's3://your_s3_path_to_the_folder_where_the_file_is_located' ;

Consulte el manual sobre la tabla Hive aquí: https://cwiki.apache.org/confluence/display/Hive/LanguageManual+DDL#LanguageManualDDL-CreateTableCreate/Drop/TruncateTable

Para ser precisos, s3 bajo el capó no almacena carpetas, el nombre de archivo que contiene /s en s3 está representado por diferentes herramientas, como Hive, como una estructura de carpetas. Ver aquí: https://stackoverflow.com/a/42877381/2700344

over 4 years ago · Santiago Trujillo Report
Answer question
Find remote jobs

Discover the new way to find a job!

Top jobs
Top job categories
Business
Post vacancy Pricing Sales
Legal
Terms and conditions Privacy policy
© 2026 PeakU Inc. All Rights Reserved.
Andres GPT
Show me some job opportunities
There's an error!