We invite you to shape the future of IBM, including product roadmaps, by submitting ideas that matter to you the most. Here's how it works:
Post your ideas
Post ideas and requests to enhance a product or service. Take a look at ideas others have posted and upvote them if they matter to you,
Post an idea
Upvote ideas that matter most to you
Get feedback from the IBM team to refine your idea
Help IBM prioritize your ideas and requests
The IBM team may need your help to refine the ideas so they may ask for more information or feedback. The product management team will then decide if they can begin working on your idea. If they can start during the next development cycle, they will put the idea on the priority list. Each team at IBM works on a different schedule, where some ideas can be implemented right away, others may be placed on a different schedule.
Receive notification on the decision
Some ideas can be implemented at IBM, while others may not fit within the development plans for the product. In either case, the team will let you know as soon as possible. In some cases, we may be able to find alternatives for ideas which cannot be implemented in a reasonable time.
The Hadoop name node GetFile Operation slowing down the job performance
While the dataset is getting generated, the segmented files with zero byte gets created first then the process start filling the data on the segmented files sequentially. If there isn’t enough volume to fill all the 8 segments ( CSDP uses 8 segments ), then the segmented files are left empty. While during the read operation, these zero byte files are getting read by Hadoop NN and cause an overhead to GetFile Operation during the peak hours. Can this design will be changed to create the segmented files only if there is a record to fill ? This will reduce the creation of zero byte files and reduce the IO overhead.
Do not place IBM confidential, company confidential, or personal information into any field.