Specific duties will include developing enhancements to our Hadoop/Hive and Redshift based data pipeline and extend to streamed data processing; contributing to data reporting projects including generating internal and external reports or visualizations (CSV, Excel, PDF, interactive web graphics and others) and data distribution / reporting components, real-time bidding or optimization algo using a combination of SQL, Ruby, R, Go, Java, Javascript, D3, Bourne shell scripts, cron jobs, and other relevant technologies. Experience with Ooozie or other Hadoop workflow solutions and experience developing complex data processing pipelines, including experience developing regressions tests and deployment strategies for such environments.