Home icon

How FINRA established real-time operational observability for Amazon EMR big data workloads on Amazon EC2 with Prometheus and Grafana

Big Data Blog



This article discusses how FINRA established real-time operational observability for their Amazon EMR big data workloads on Amazon EC2 using Prometheus and Grafana.

Specifically, the article covers:

  • The challenges FINRA faced in monitoring their complex and dynamic EMR clusters, such as scale, data variety, resource utilization, latency metrics, centralized dashboards, alerting, and cost management.
  • The solution architecture using Amazon Managed Prometheus for metric collection, Amazon Managed Grafana for visualization dashboards, and Amazon CloudWatch for alerting.
  • Details on the customized Grafana dashboards displaying node, cluster, and job-level metrics like OS, HDFS, YARN, Spark, and JVM.
  • The conclusion highlighting how this solution enabled comprehensive observability, performance optimization, and operational excellence for FINRA's big data workloads.


Go to article

The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.

Related articles

Nov 15
2024
Amazon CloudWatch launches Observability Solutions for AWS Services and Workloads on AWS
Sep 27
2024
Amazon EMR Serverless observability, Part 1: Monitor Amazon EMR Serverless workers in near real time using Amazon CloudWatch
Oct 29
2024
Governing the ML lifecycle at scale: Centralized observability with Amazon SageMaker and Amazon CloudWatch
Nov 20
2024
Monitor EBS Detailed Performance Statistics with Amazon Managed Service for Prometheus

The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.