Loading...

Messages

Proposals

Stuck in your homework and missing deadline? Get urgent help in $10/Page with 24 hours deadline

Get Urgent Writing Help In Your Essays, Assignments, Homeworks, Dissertation, Thesis Or Coursework & Achieve A+ Grades.

Privacy Guaranteed - 100% Plagiarism Free Writing - Free Turnitin Report - Professional And Experienced Writers - 24/7 Online Support

Big data processing rmit

28/03/2021 Client: saad24vbs Deadline: 2 Day

RMIT Classification: Trusted

Big Data Processing COSC 2637/2633

Assignment 1 Assessment Type

Individual assignment. Submit online via Canvas → Assignment 1. Marks awarded for meeting requirements as closely as possible. Clarifications/updates may be made via announcements or relevant discussion forums.

Due Date Week 7, Friday 13rd September 2020, 23:59

Marks 40

1. Overview

Write MapReduce programs which gives your chance to understand the complexity of MapReduce programing, the essential components you learned in lectures, the unique debugging method, the impact of performance using different size clusters.

2. Learning Outcomes The key course learning outcomes are:

CLO 1. Model and implement efficient big data solutions for various application areas using appropriately selected algorithms and data structures.

CLO 2. Analyze methods and algorithms, to compare and evaluate them with respect to time and space requirements and make appropriate design choices when solving real-world problems.

CLO 3. Motivate and explain trade-offs in big data processing technique design and analysis in written and oral form.

CLO 4. Explain the Big Data Fundamentals, including the evolution of Big Data, the characteristics of Big Data and the challenges introduced.

CLO 5. Apply non-relational databases, the techniques for storing and processing large volumes of structured and unstructured data, as well as streaming data.

CLO 6. Apply the novel architectures and platforms introduced for Big data, in particular Hadoop and MapReduce.

3. Assessment details In Task 2 of Lab 3 (week 4), you have developed a MapReduce program and run it in Hadoop. It is the basic version of word count. In this assignment, you are asked to extend the functions based on the MapReduce program using what you learned in this course. You should use Java to develop your MapReduce program over AWS EMR (if you want to use other code language, please contact lecturer for approval). Task 1 – Count words by lengths (8 marks) Write a MapReduce program to count number of short words (1-4 letters), medium words (5-7 letters) words, long words (8-10 letters) and extra-long words (More than 10 letters). Task 2 – Count words by the first character (8 marks) Write a MapReduce program that outputs a count of all words that begin with a vowel and count of all how many words that begin with a consonant.

Page 2 of 3

RMIT Classification: Trusted

Task 3 – Count word with in-mapper combining (12 marks) Write a MapReduce program to count the number of each word where the in-mapper combining is implemented rather than an independent combiner. Task 4 – Count word with partitioner (12 marks) Extend the MapReduce code in Task 1 by using partitioner such that - short words (1-4 letters) and extra-long words (More than 10 letters) are processed in one reducer, - medium words (5-7 letters) and long words (8-10 letters) are processed in another reducer.

4. Submission Your assignment should follow the requirement below and submit via Canvas > Assignment 1. Assessment declaration: when you submit work electronically, you agree to the assessment declaration: https://www.rmit.edu.au/students/student-essentials/assessment-and-exams/assessment/assessment-declaration

5. Requirement

(a) The codes for all four tasks are entailed in a single Maven project. (2 marks) (b) Submit the complete Maven project source code in a .zip file (including a standalone jar file). The zip

file should be named as sxxxxx_BDP_S2_2020.zip (replace sxxxxx by your student ID). (2 marks) (c) You need include a “README” file in the zip file. In the README, you are asked to specify how

to run each task using the standalone jar in Hadoop. (1 mark) (d) Paths of input file and output file should not be hard-coded. (1 mark) (e) For all tasks, use the text file “Melbourne” “RMIT” and “3littlepigs” as the input files and process

them together in the same run (don’t process them separately); output file must be stored in /user/sxxxxx/output# in HDFS (i.e., /user/sxxxxx/output1 for Task 1, /user/sxxxxx/output2 for Task 2, and so on). (4x1 marks)

(f) For each task, using Apache log4j log information: (4x1 marks) - In the MAP tasks, the log should be “The mapper task of , ” - In the REDUCE tasks, the log should be “The reducer task of ,

(g) Conduct performance analysis on different numbers of nodes in EMR clusters. To this end, run the code in Task 1 to process a large data set

s3a://commoncrawl/crawl-data/CC-MAIN-2018-17/segments/1524125936833.6 when the number of nodes in EMR clusters is 3, 5, 7 respectively. Show the CPU_MILLISECONDS for each MAP task and each REDUCE task in the README file (the same one mentioned in (c)); and analyze what you observed (250-500 words). (3 marks)

(h) Your MapReduce program(s) must be well written, using good coding style and including appropriate use of comments. (4x2 marks)

6. Marking Guide (a) If one task cannot be run using the submitted jar file, no mark for this task. (b) If one task can run but the output is incorrect. At least half mark will be deducted for this task. If the

code has major issues (such as logically incorrect), 0 mark for this task.

7. Academic integrity and plagiarism (standard warning) Academic integrity is about honest presentation of your academic work. It means acknowledging the work of others while developing your own insights, knowledge and ideas. You should take extreme care that you have:

Homework is Completed By:

Writer Writer Name Amount Client Comments & Rating
Instant Homework Helper

ONLINE

Instant Homework Helper

$36

She helped me in last minute in a very reasonable price. She is a lifesaver, I got A+ grade in my homework, I will surely hire her again for my next assignments, Thumbs Up!

Order & Get This Solution Within 3 Hours in $25/Page

Custom Original Solution And Get A+ Grades

  • 100% Plagiarism Free
  • Proper APA/MLA/Harvard Referencing
  • Delivery in 3 Hours After Placing Order
  • Free Turnitin Report
  • Unlimited Revisions
  • Privacy Guaranteed

Order & Get This Solution Within 6 Hours in $20/Page

Custom Original Solution And Get A+ Grades

  • 100% Plagiarism Free
  • Proper APA/MLA/Harvard Referencing
  • Delivery in 6 Hours After Placing Order
  • Free Turnitin Report
  • Unlimited Revisions
  • Privacy Guaranteed

Order & Get This Solution Within 12 Hours in $15/Page

Custom Original Solution And Get A+ Grades

  • 100% Plagiarism Free
  • Proper APA/MLA/Harvard Referencing
  • Delivery in 12 Hours After Placing Order
  • Free Turnitin Report
  • Unlimited Revisions
  • Privacy Guaranteed

6 writers have sent their proposals to do this homework:

Top Academic Tutor
ECFX Market
Essay & Assignment Help
Academic Mentor
Quality Homework Helper
Top Quality Assignments
Writer Writer Name Offer Chat
Top Academic Tutor

ONLINE

Top Academic Tutor

Give me a chance, i will do this with my best efforts

$63 Chat With Writer
ECFX Market

ONLINE

ECFX Market

I am known as Unrivaled Quality, Written to Standard, providing Plagiarism-free woork, and Always on Time

$44 Chat With Writer
Essay & Assignment Help

ONLINE

Essay & Assignment Help

I have read and understood all your initial requirements, and I am very professional in this task.

$64 Chat With Writer
Academic Mentor

ONLINE

Academic Mentor

I will cover all the points which you have mentioned in your project details.

$58 Chat With Writer
Quality Homework Helper

ONLINE

Quality Homework Helper

You can award me any time as I am ready to start your project curiously. Waiting for your positive response. Thank you!

$53 Chat With Writer
Top Quality Assignments

ONLINE

Top Quality Assignments

I have read and understood all your initial requirements, and I am very professional in this task.

$36 Chat With Writer

Let our expert academic writers to help you in achieving a+ grades in your homework, assignment, quiz or exam.

Similar Homework Questions

A survey of teenagers 12–17 indicated three circumstances in which they were more likely to use drugs: - Hard rock cafe operations management in services case study - Aspen plastics produces plastic bottles - Nursing practice - Ip http server packet tracer - Pearl e white orthodontist specializes in correcting misaligned teeth - Wd my cloud unable to access device 503 - The cyclotron mastering physics - Does katniss love peeta - Colleagues Response - Do not resuscitate ethical issues - Assignment 2 project paper comparative essay - 13th documentary discussion questions - Disadvantages of selective breeding - What is the difference between volume and capacity - Ids 100 project 1 lenses chart - Plant Physiology - Cis 9 status report - Risks that bill gates took - Final Paper part 1 due in 30 hours - Hydrochloric acid 0.1 n msds - Apply a manual summary route for the loopback interface networks - Ocr a level biology multiple choice questions - Marie winn tv addiction pdf - Managerial accounting and cost concepts ppt - Jasper jones analysis pdf - Compare the two anaerobic energy systems - Muslim Molvi 7340613399 OnLine No 1 FaMOUs VashIKaraN sPecIaLIsT IN Bhatpara - Getting paid math worksheet 2.3 9 a1 answers - Question - Oxford island nature reserve - What are the first five multiples of 13 - Belgium beer dan murphy - Boq specialist qantas points - Question - Ethane reacts with bromine in the presence of ultraviolet light - English Composition - Security assessment report sar template - Fisher price company information - Queensland health annual refresher - Information technology in Global Economy - Accommodations for diverse learners - What is the chemical formula for aluminum bromide - Guys and dolls brantham - Offred acts of rebellion - San diego mesa college tutoring appointment - No1 ccytotec pills +27835179056 NEW HOPE WOMEN ABORTION CLINIC in St Michael's-on-Sea Umkomaas Umtentweni Umzinto - Quick answer - Https www youtube com watch v 7majoi3qmu0 - International human resource management a multinational company perspective - Terrorism and organized crime - Energy value of foods table - Fltv10010 - making movies 1 - What is green marketing myopia - Completing the square worksheet - George kyparisis - Governments often intervene in international trade and impose quotas to - GBI Assignment - Literal and implied meaning worksheets - Working capital simulation managing growth v2 answers - 1993 big bayou canot train wreck victims - Cisco callmanager attendant console - Hoselink 15m retractable hose reel - AIS - Assignment 2 - The unadjusted trial balance of epicenter laundry - Week 2 Project - Watkins silentflo 5000 making noise - Reflection 7 - Master budget template excel - 50-C2 - Large heavy protective glove crossword clue - What is international marketing task - Is dramatic irony a dramatic technique - The learning odyssey answers economics - Functions of endocrine gland hormones matching - Man and his symbols review - Anagrams of christmas carols and songs - According to the rokeach value survey - Discussion on Emerging threats - Pte brisbane study centre - Business information system assignment - REPLY 1TPofN - Standard addition method equation - Weiss functional impairment rating scale for adults - Preventing injury gcse pe - Ielts speaking band descriptors - Criminal Justice/discussion - Journal article analysis - Why does cryptography software fail - The new chinese astrology by suzanne white pdf - Ethical and Legal Implications of Prescribing Drugs - Rockwell collins mpe s - Jitterbugs dancing the chemical reaction steps answers - Personal identity - Electric field created by a point charge - Effective business writing quiz - ORGAN THEORY DQ 5# - What is the difference between job analysis and job design - Sorensen systems inc is expected to pay