The document describes the CoHadoop system, which extends Hadoop to enable flexible data placement by colocating related files. CoHadoop introduces the concept of a locator, where files with the same locator will be stored on the same set of data nodes. This improves performance for operations like joins and indexes that work on related files by allowing them to access data locally without network overhead. Experimental results show that CoHadoop provides up to 60% faster performance for join queries compared to the default Hadoop data placement.