The leading knowledge platform for the financial technology industry
The leading knowledge platform for the financial technology industry

A-Team Insight Blogs

ScaleOut Pushes Hadoop Towards Low-Latency for Real-Time Analytics

OK, so the headline is a tad extreme, but bear with me. Recent developments combining in-memory technologies and Hadoop/MapReduce from ScaleOut Software point to a future where big data analytics and real-time processing, as it’s defined in the financial markets, could meet.

ScaleOut has just released its ScaleOut hServer V2, an in-memory data grid, which it claims can boost Hadoop performance by 20x, and can make it suitable for processing ‘live data’ to deliver ‘rea-ltime analytics’.

“To minimise execution time, ScaleOut hServer employs numerous optimisations to minimise data motion during the execution of MapReduce applications, and it can automatically cache HDFS data sets within the IMDG (a feature introduced with ScaleOut hServer V1). In addition, ScaleOut hServer’s memory capacity and throughput can be scaled by adding servers to the IMDG’s cluster. The product automatically rebalances the data set and execution workload when servers are added or removed,” says the company in a statement.

As well as boosting performance of a Hadoop deployment, hServer also incorporates Map/Reduce logic so that a Hadoop distribution is not actually required – though the company suggests its offering is not a direct replacement for Hadoop.

Nevertheless, “ScaleOut hServer is designed to be compatible with most Java-based Hadoop Map/Reduce applications developed for the standard Hadoop distributions, requiring only a one-line code change to execute applications using ScaleOut hServer.”

The big picture here is that ScaleOut – as well as other companies pushing in-memory technology – is recognising that the batch-oriented nature of Hadoop has limitations for real-time applications, such as those found in the financial markets.

While ScaleOut is today looking to boost Hadoop performance to make applications that used to take hours and minutes to execute run now in minutes and seconds, the performance trajectory could well follow that of the low-latency space, where milliseconds gave way to microseconds, and now nanoseconds.

The deployment of multi-core and multi-socket servers, GPU technologies and advances in memory will all benefit data grid vendors like ScaleOut, as well as Hadoop and other big data analytics offerings.

Related content

WEBINAR

Upcoming Webinar: Infrastructure monitoring: mapping technical performance to business performance

Date: 8 July 2021 Time: 10:00am ET / 3:00pm London / 4:00pm CET Duration: 50 minutes It’s a widely recognized truth that if you can’t measure something, you can’t improve its performance. As high-performance connectivity technologies have established themselves in the mainstream of financial firms’ trading architectures, the ability to monitor messaging and data infrastructures...

BLOG

London’s Continued Future as a Global Financial Centre Assured in Post-Brexit Era

By John R. Bryson, Professor of Enterprise and Economic Geography, University of Birmingham Recent newspaper headlines have declared that ‘Amsterdam ousts London as Europe’s top share trading hub’. This has been seen as another downside of the UK leaving the UK. Amsterdam has become Europe’s most important hub for trading shares with London pushed into second place....

EVENT

RegTech Summit New York City

Now in its 5th year, the RegTech Summit in NYC explores how the North American financial services industry can leverage technology to drive innovation, cut costs and support regulatory change.

GUIDE

The Data Management Challenges of Client Onboarding and KYC

This special report accompanies a webinar we held on the popular topic of The Data Management Challenges of Client Onboarding and KYC, discussing the data management challenges of client onboarding and KYC, and detailing new technology solutions that have the potential to automate and streamline onboarding and KYC processes. You can register here to get immediate...