Setting Up a Git Server Using Gitosis

Posted on

Update: Since gitosis is not maintained and supported, please check out gitolite for setting up a new git server. (see the comment from Sitaram Chamarty, the gitolite author, the author of gitolite.) Gitosis is a piece of software writen by Tommi Virtanen for hosting git repositories. It manages multiple repositories under the same user account.
Read more

Hadoop Default Ports

Posted on

Hadoop’s namenode and datanodes expose a bunch of TCP ports used by Hadoop’s daemons to communicate to each other or listen directly to users’ requests. These ports information are needed by both the Hadoop users and cluster administrators to write programs or configure firewalls/gateways accordingly. A post written by Philip Zeyliger from Cloudera’s blog summarizes the
Read more

A Simple Sort Benchmark on Hadoop

Posted on

After [[hadoop-installation-tutorial|installing Hadoop]], we usually run some benchmark programs to test whether the system works well. In the post of the Hadoop install tutorial, we show a very simple to grep strings from a simple sets of files. In this post, we introduce the Sort for testing and benchmarking Hadoop. The Sort program is also
Read more

Conferences on Cloud Computing 2012

Posted on

This post lists important conferences on Cloud Computing in year 2012. OSDI 2012 Table of Contents OSDI 2012EuroSys 2012USENIX ATC ’12ASPLOS 2012VEE 2012NSDI 2012HotPar ’12SOCC 2012 10th USENIX Symposium on Operating Systems Design and Implementation (OSDI ’12) October 8–10, 2012, Hollywood, CA “The tenth OSDI seeks to present innovative, exciting research in computer systems. OSDI
Read more

Pitfalls and Lessons on Configuing and Tuning Hadoop

Posted on

This post lists pitfalls and lessons learning when configuring and tuning Hadoop. Hadoop with IPv6 Table of Contents Hadoop with IPv6Hostname vs. IPJava Virtual Machine Hadoo doesn’t support IPv6 currently (up to 0.20.2 and 0.21.0): Hadoop and IPv6. The performance of the cluster may suffer from turning IPv6 on in clusters: mail archive. One good
Read more

Setting Up Standalone (Local) Hadoop

Posted on

Hadoop is designed to run on [[hadoop-installation-tutorial|hundreds to thousands of computers]] inside cluster. However, Hadoop is configured to run things in a non-distributed mode as a single Java process by default. This is specially useful for debugging since distributed debugging is really a nightmare. This post introduces how to set up a standalone Hadoop environment.
Read more

Conferences on Cloud Computing 2011

Posted on

This post lists important conferences on Cloud Computing in year 2011. ACM Symposium on Cloud Computing Table of Contents ACM Symposium on Cloud Computing23rd ACM Symposium on Operating Systems Principles (SOSP)EuroSys 2011CLOUD COMPUTING 2011 October 27 and 28, 2011, Cascais, Portugal Submission Deadline: April 30, 2011 23rd ACM Symposium on Operating Systems Principles (SOSP) October
Read more

mrcc – A Distributed C Compiler System on MapReduce

Posted on

The mrcc project’s homepage is here: mrcc project. Abstract Table of Contents AbstractIntroductionDesign and architectureImplementationmrccmrcc-mapEvaluationInstallation and configurationInstall hadoopConfigurationCompile and install:How to use itCompile the project using mrcc on MapReduceEnable detailed logging if neededRelated workReferences mrcc is an open source compilation system that uses MapReduce to distribute C code compilation across the servers of the cloud
Read more