mrjob

_images/logo_medium.png

mrjob is a Python package that helps you write and run Hadoop Streaming jobs.

mrjob fully supports Amazon’s Elastic MapReduce (EMR) service, which allows you to buy time on a Hadoop cluster on an hourly basis. It also works with your own Hadoop cluster.

Table of Contents

Indices and tables

Table Of Contents

Next topic

What’s New

This Page