We are looking for a highly skilled software engineer to build next-generation AI data platforms and lakehouse solutions. The ideal candidate will have deep expertise in database technologies, modern data warehouses, open table formats, distributed data
Candidate should be able to: Coordinate Development, Integration, and Production deployments. Optimize Spark code, Impala queries, and Hive partitioning strategy for better scalability, reliability, and performance. Build applications using Maven, SBT and integrated with continuous integration
• Developing the Best practices documentation to the Application Developer partners to educate how to use the Cloudera Services. • install/Configure/Maintain Apache Cloudera tools like Hive, Impala, Sqoop, Flume, Kafka, HBase, SOLR, and File formats like Avro, Parquet.