Java IoT Authors: Liz McMillan, John Mertic, Yeshim Deniz, Carmen Gonzalez, Elizabeth White

Related Topics: Java IoT

Java IoT: Article

Java Performance I/O Tuning

Java Performance I/O Tuning

Many Java programs that utilize I/O are excellent candidates for performance tuning. One of the more common problems in Java applications is inefficient I/O. A profile of Java applications and applets that handle significant volumes of data will show significant time spent in I/O routines, implying substantial gains can be had from I/O performance tuning. In fact, the I/O performance issues usually overshadow all other performance issues, making them the first area to concentrate on when tuning performance. Therefore, I/O efficiency should be a high priority for developers looking to optimally increase performance. Unfortunately, optimal reading and writing can be challenging in Java.

Once an application's reliance upon I/O is established and I/O is determined to account for a substantial slice of the applications execution time, performance tuning can be undertaken. The best method for determining the distribution of execution time among methods is to use a profiler. Sunª Javaª WorkShopª software provides an excellent profiler that offers detailed call counts and execution times for each method. System method call statistics can be tabulated as an option. Stream chaining and custom I/O class methods of performance tuning are discussed. An example program is provided that allows the progressive measurement of the progress of the tuning effort. Using the example program that is provided, JavaIOTest.java, and utilizing the techniques described, substantial performance improvements of an order of magnitude can be achieved. Simple stream chaining provides approximately a 91% decrease in execution time from 28,198 milliseconds to 2,510 milliseconds, while a custom BufferedFileReader class cuts performance time by another 75%, over 97% total, to 630 milliseconds for a 250 kilobyte text file on The Sun™ Solaris™ 2.6 operating environment.

Java performance is currently a topic of great interest. Performance is usually hotly debated for any relatively new language or operating environment, so this is not surprising. However, Java's reliance upon the availability of sufficient network bandwidth for the downloading of classes shifts the relative benefits of some options for optimization. The reliance on the network penalizes optimization techniques that favor increasing code size in order to provide faster execution. The resulting optimized classes can take longer to download to the client. Of course, server-side Java is not as acutely affected by code size and developers can even consider native code compilers for that case. Based upon anecdotal evidence, most Java development today seems to be concentrated on client-side applets with the result that download times are an important criterion. Java optimization efforts, therefore, need to be well-researched and considered.

Because Java is a relatively new language, optimizing compiler features are less sophisticated than those available for C and C++, leaving room for more "hand-crafting". The "hand" optimization of key sections identified by profilers, such as the profiler available in Sun's Java WorkShop 2.0, can reap substantial benefits.

One of the more common problems in Java applications is inefficient I/O. A profile of Java applications and applets that handle significant volumes of data will show significant time spent in I/O routines, implying substantial gains can be had from I/O performance tuning. In fact, the I/O performance issues usually overshadow all other performance issues, making them the first area to concentrate on when tuning performance. Therefore, I/O efficiency should be a high priority for developers looking to optimally increase performance. Unfortunately, optimal reading and writing can be challenging in Java. Streamlining the use of I/O often results in greater performance gains than all other possible optimizations combined. It is not uncommon to see a speed improvement of at least an order of magnitude using efficient I/O techniques, as this paper and the example program will demonstrate.

This article focuses on the improvement gains possible through careful use of both the existing Java I/O classes and the introduction of a custom file reader, BufferedFileReader. BufferedFileReader is responsible for some of the performance increase of Java WorkShop version 2.0 over version 1.0. An example application is used to read three different file sizes, ranging from 100 kilobytes to 500 kilobytes and the results are compared for various optimizations.

Performance Tuning Through Stream Chaining
As a demonstration of I/O performance tuning, this article will describe the process of tuning a sample program created expressly for this paper: JavaIOTest. JavaIOTest tracks the execution times for several I/O schemes starting with a very basic DataInputStream method and culminating with the use of a custom-buffered, file-reader class, while demonstrating the performance improvements obtained by several program design changes during the tuning effort. The actual execution times are meant to show the relative improvements possible.* The actual execution times will vary widely among the systems used. Readers are cautioned that what is important is the relative improvement on the same system, test-to-test, and that comparisons across operating environments and systems are complex and the results can be specious.

Basic IO: DataInputStream
The I/O method used in this section is a DataInputStream chained to a FileInputStream as shown in Listing 1. This method of reading a file is very common since it is simple, but it is extremely slow. The reason for the poor performance is that the DataInputStream class does no buffering. The resulting reads are done one byte at a time. Several instances of this technique have been found in the JDKª software as well as several "real" Java programs, providing fertile ground for improvement through a tuning regime (see Listing 1).

The results of using the default, basic I/O scheme are as follows. The first section of the example program, JavaIOTest, showed run times of 28,198 milliseconds reading a 250 kilobyte file.*

An Improvement: BufferedInputStream
A simple improvement involves buffering the FileInputStream by interposing a BufferedInputStream in the stream chain. This buffers the data, with the default buffer size of 2048 bytes. Listing 2 illustrates the minor source code change required.

The resulting performance increase for the medium sized file (250 kilobytes) was 91%, from 28,198 milliseconds to 2,510 -- over an order of magnitude with just a simple change.*

The New JDK 1.1 Classes
The foregoing method has provided a substantial performance improvement but has a serious flaw: the readLine() method of DataInputStream does not properly handle Unicode characters. The problem is that the method assumes all characters are one byte in length while Unicode characters are two bytes in length. This method has been deprecated beginning in JDK 1.1. Since deprecated classes are discouraged, the FileReader and BufferedReader classes should be substituted for the classes.

Unfortunately, the scheme to provide for Unicode character localization consists of invoking a locale-dependent converter on the raw bytes to convert them to Java characters, causing an extra copy operation per character. This penalty is offset by other efficiencies in the code. The code change is shown in Listing 3.

The resulting performance increase for the medium file size was 57%, >from 2,510 to 1,092 milliseconds.*

Buffer Size Effects
The buffer size used in buffering schemes is important for performance. As a rule of thumb, bigger is better to a point. In order to examine the impact of the buffer size, a test run was made with a smaller buffer than the default of 8,192 bytes used in the BufferedReader class. Listing 4 shows the code segment using a reduced buffer size of 1,024 bytes.

Depending upon the file's size and platform used for testing, the larger buffer size provided performance improvements ranging from 3 to 13 percent. The use of a large buffer size will improve performance significantly and should be considered unless local memory is restricted.

Using simple stream-chaining techniques, the execution performance of an I/O bound Java program has been increased an average of 97 percent over using the simple DataInputStream class. This is a substantial improvement for a little extra design work and one that could mean the difference between shipping and re-designing an interactive application.

Tuning with Custom I/O Classes
To this point, tuning has focused on using the core classes distributed with the JDK. With each version of the JDK, more effort seems to be going into tuning critical sections for performance. The improvement in speed of the BufferedReader class over the BufferedInputStream class despite the additional copy per character hints at this. However, if the application needs to read large files, a custom class can be created to further tune performance. The BufferedReader.readLine() method creates an instance of StringBuffer to hold the characters in the line it reads. It then converts the StringBuffer to String, resulting in two more copies per character. The BufferedFileReader class utilizes a modified readLine() method that avoids the extra, double-copy in most cases. It also adds the convenience of creating the FileReader class for the caller. Listing 5 shows the changes required to use this class. The resulting performance increase for the medium file size was 32% overall to 630 milliseconds.*

The BufferedFileReader class is being used in Java WorkShop (package sun.jws.util). The documentation comment in Listing 6 describes the efficiencies added.

Without having to chain together several different classes, as with the standard JDK classes, the example provides a single, efficient class through which a file may be read. It is also more efficient (typically faster) than the fastest JDK classes. Specific optimizations include:
1. More efficiently coded readLine( ) method.
2. Adds open( ) method, so the class can be reused when several files are read in a loop. This avoids repeated allocation and deallocation of buffers.

This class (see Listing 6) contains a self-benchmarking test in its main() method that can be used to measure the exact speedup on a particular system.

Further Tuning
Although the example in Listing 5 is as much as 45 times faster than the example in Listing 1 (and actually comprises fewer lines of code), it is still far from the best that can be done. There are at least two more major optimizations that can be done if still higher performance is required and we are willing to do a little more work.

First, if we look at the first line of the while loop, we see that a new String object is being created for every line of the file being read:

while ((line = in.readLine()) != null) {

This means, for example, that for a 100,000 line file 100,000 String objects would be created. Creating a large number of objects incurs costs in three ways: 1. Time and memory to allocate the space for the objects
2. Time to initialize the objects
3. Time to garbage collect the objects

The problem here is that the I/O buffer is private; the user cannot access it directly. Therefore, BufferedFileReader must create a new String object in order to return the data to the user. Although this follows the conventional assertion that class structures should largely be private in order to control data access, the performance penalty is too high an insurance premium for this case.

To get around this problem, the user must manage the buffer directly without using the BufferedReader or BufferedFileReader convenience classes. This will enable the user to reuse buffers rather than creating a new object each time to hold the data.

Second, strings are inherently less efficient than arrays based upon char. This is because the user must call a method to access each character of a String, whereas the characters can be accessed directly in a char array. Hence, our code example can be made more efficient by avoiding Strings entirely, and using char arrays directly.

Listing 7 shows the code which implements the two optimizations above. It is substantially more lines of code than the previous examples, but tests show it performs as much as 3 times faster than the example in Listing 5.

Performance Tuning Results
The results of running the test program used for this article, JavaIOTest, on text files ranging from 100 kilobytes to 500 kilobytes in size are summarized in the tables found in Appendix One at http://www.sun.com/workshop/java/wp-javaio. The relative performance numbers are more important than the absolute numbers since the system was not isolated nor used exclusively for just the test processes. As the automobile industry states in its disclaimers, "your mileage may vary", readers are again cautioned that what is most important is the relative improvement on the same system and test-to-test. Comparisons across operating environments and systems are complex and the results can be specious.

*Full test results are available at: http://www.sun.com/workshop/java/wp-javaio. See Appendix One.

This article was provided by Engineering, Sun's Authoring and Development Tools Group.

Comments (0)

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.

@ThingsExpo Stories
SYS-CON Events announced today that Roundee / LinearHub will exhibit at the WebRTC Summit at @ThingsExpo, which will take place on November 1–3, 2016, at the Santa Clara Convention Center in Santa Clara, CA. LinearHub provides Roundee Service, a smart platform for enterprise video conferencing with enhanced features such as automatic recording and transcription service. Slack users can integrate Roundee to their team via Slack’s App Directory, and '/roundee' command lets your video conference ...
SYS-CON Events announced today that Enzu will exhibit at the 19th International Cloud Expo, which will take place on November 1–3, 2016, at the Santa Clara Convention Center in Santa Clara, CA. Enzu’s mission is to be the leading provider of enterprise cloud solutions worldwide. Enzu enables online businesses to use its IT infrastructure to their competitive advantage. By offering a suite of proven hosting and management services, Enzu wants companies to focus on the core of their online busine...
SYS-CON Events announced today that SoftNet Solutions will exhibit at the 19th International Cloud Expo, which will take place on November 1–3, 2016, at the Santa Clara Convention Center in Santa Clara, CA. SoftNet Solutions specializes in Enterprise Solutions for Hadoop and Big Data. It offers customers the most open, robust, and value-conscious portfolio of solutions, services, and tools for the shortest route to success with Big Data. The unique differentiator is the ability to architect and...
In past @ThingsExpo presentations, Joseph di Paolantonio has explored how various Internet of Things (IoT) and data management and analytics (DMA) solution spaces will come together as sensor analytics ecosystems. This year, in his session at @ThingsExpo, Joseph di Paolantonio from DataArchon, will be adding the numerous Transportation areas, from autonomous vehicles to “Uber for containers.” While IoT data in any one area of Transportation will have a huge impact in that area, combining senso...
Why do your mobile transformations need to happen today? Mobile is the strategy that enterprise transformation centers on to drive customer engagement. In his general session at @ThingsExpo, Roger Woods, Director, Mobile Product & Strategy – Adobe Marketing Cloud, covered key IoT and mobile trends that are forcing mobile transformation, key components of a solid mobile strategy and explored how brands are effectively driving mobile change throughout the enterprise.
@ThingsExpo has been named the Top 5 Most Influential Internet of Things Brand by Onalytica in the ‘The Internet of Things Landscape 2015: Top 100 Individuals and Brands.' Onalytica analyzed Twitter conversations around the #IoT debate to uncover the most influential brands and individuals driving the conversation. Onalytica captured data from 56,224 users. The PageRank based methodology they use to extract influencers on a particular topic (tweets mentioning #InternetofThings or #IoT in this ...
"Matrix is an ambitious open standard and implementation that's set up to break down the fragmentation problems that exist in IP messaging and VoIP communication," explained John Woolf, Technical Evangelist at Matrix, in this SYS-CON.tv interview at @ThingsExpo, held Nov 4–6, 2014, at the Santa Clara Convention Center in Santa Clara, CA.
The IoT has the potential to create a renaissance of manufacturing in the US and elsewhere. In his session at 18th Cloud Expo, Florent Solt, CTO and chief architect of Netvibes, discussed how the expected exponential increase in the amount of data that will be processed, transported, stored, and accessed means there will be a huge demand for smart technologies to deliver it. Florent Solt is the CTO and chief architect of Netvibes. Prior to joining Netvibes in 2007, he co-founded Rift Technologi...
For basic one-to-one voice or video calling solutions, WebRTC has proven to be a very powerful technology. Although WebRTC’s core functionality is to provide secure, real-time p2p media streaming, leveraging native platform features and server-side components brings up new communication capabilities for web and native mobile applications, allowing for advanced multi-user use cases such as video broadcasting, conferencing, and media recording.
Established in 1998, Calsoft is a leading software product engineering Services Company specializing in Storage, Networking, Virtualization and Cloud business verticals. Calsoft provides End-to-End Product Development, Quality Assurance Sustenance, Solution Engineering and Professional Services expertise to assist customers in achieving their product development and business goals. The company's deep domain knowledge of Storage, Virtualization, Networking and Cloud verticals helps in delivering ...
24Notion is full-service global creative digital marketing, technology and lifestyle agency that combines strategic ideas with customized tactical execution. With a broad understand of the art of traditional marketing, new media, communications and social influence, 24Notion uniquely understands how to connect your brand strategy with the right consumer. 24Notion ranked #12 on Corporate Social Responsibility - Book of List.
More and more brands have jumped on the IoT bandwagon. We have an excess of wearables – activity trackers, smartwatches, smart glasses and sneakers, and more that track seemingly endless datapoints. However, most consumers have no idea what “IoT” means. Creating more wearables that track data shouldn't be the aim of brands; delivering meaningful, tangible relevance to their users should be. We're in a period in which the IoT pendulum is still swinging. Initially, it swung toward "smart for smar...
Cognitive Computing is becoming the foundation for a new generation of solutions that have the potential to transform business. Unlike traditional approaches to building solutions, a cognitive computing approach allows the data to help determine the way applications are designed. This contrasts with conventional software development that begins with defining logic based on the current way a business operates. In her session at 18th Cloud Expo, Judith S. Hurwitz, President and CEO of Hurwitz & ...
@ThingsExpo has been named the Top 5 Most Influential M2M Brand by Onalytica in the ‘Machine to Machine: Top 100 Influencers and Brands.' Onalytica analyzed the online debate on M2M by looking at over 85,000 tweets to provide the most influential individuals and brands that drive the discussion. According to Onalytica the "analysis showed a very engaged community with a lot of interactive tweets. The M2M discussion seems to be more fragmented and driven by some of the major brands present in the...
In the next five to ten years, millions, if not billions of things will become smarter. This smartness goes beyond connected things in our homes like the fridge, thermostat and fancy lighting, and into heavily regulated industries including aerospace, pharmaceutical/medical devices and energy. “Smartness” will embed itself within individual products that are part of our daily lives. We will engage with smart products - learning from them, informing them, and communicating with them. Smart produc...
As ridesharing competitors and enhanced services increase, notable changes are occurring in the transportation model. Despite the cost-effective means and flexibility of ridesharing, both drivers and users will need to be aware of the connected environment and how it will impact the ridesharing experience. In his session at @ThingsExpo, Timothy Evavold, Executive Director Automotive at Covisint, will discuss key challenges and solutions to powering a ride sharing and/or multimodal model in the a...
In his keynote at 19th Cloud Expo, Sheng Liang, co-founder and CEO of Rancher Labs, will discuss the technological advances and new business opportunities created by the rapid adoption of containers. With the success of Amazon Web Services (AWS) and various open source technologies used to build private clouds, cloud computing has become an essential component of IT strategy. However, users continue to face challenges in implementing clouds, as older technologies evolve and newer ones like Docke...
SYS-CON Events announced today that Embotics, the cloud automation company, will exhibit at the 19th International Cloud Expo, which will take place on November 1–3, 2016, at the Santa Clara Convention Center in Santa Clara, CA. Embotics is the cloud automation company for IT organizations and service providers that need to improve provisioning or enable self-service capabilities. With a relentless focus on delivering a premier user experience and unmatched customer support, Embotics is the fas...
Just over a week ago I received a long and loud sustained applause for a presentation I delivered at this year’s Cloud Expo in Santa Clara. I was extremely pleased with the turnout and had some very good conversations with many of the attendees. Over the next few days I had many more meaningful conversations and was not only happy with the results but also learned a few new things. Here is everything I learned in those three days distilled into three short points.
SYS-CON Events announced today that Coalfire will exhibit at the 19th International Cloud Expo, which will take place on November 1–3, 2016, at the Santa Clara Convention Center in Santa Clara, CA. Coalfire is the trusted leader in cybersecurity risk management and compliance services. Coalfire integrates advisory and technical assessments and recommendations to the corporate directors, executives, boards, and IT organizations for global brands and organizations in the technology, cloud, health...