官术网_书友最值得收藏!

Features of Storm

The following are some of the features of Storm that make it a perfect solution to process streams of data in real time:

  • Fast: Storm has been reported to process up to 1 million tuples/records per second per node.
  • Horizontally scalable: Being fast is a necessary feature to build a high volume/velocity data processing platform, but a single node will have an upper limit on the number of events that it can process per second. A node represents a single machine in your setup that executes Storm applications. Storm, being a distributed platform, allows you to add more nodes to your Storm cluster and increase the processing capacity of your application. Also, it is linearly scalable, which means that you can double the processing capacity by doubling the nodes.
  • Fault tolerant: Units of work are executed by worker processes in a Storm cluster. When a worker dies, Storm will restart that worker, and if the node on which the worker is running dies, Storm will restart that worker on some other node in the cluster. This feature will be covered in more detail in Chapter 3, Storm Parallelism and Data Partitioning.
  • Guaranteed data processing: Storm provides strong guarantees that each message entering a Storm process will be processed at least once. In the event of failures, Storm will replay the lost tuples/records. Also, it can be configured so that each message will be processed only once.
  • Easy to operate: Storm is simple to deploy and manage. Once the cluster is deployed, it requires little maintenance.
  • Programming language agnostic: Even though the Storm platform runs on Java virtual machine (JVM), the applications that run over it can be written in any programming language that can read and write to standard input and output streams.
主站蜘蛛池模板: 秦皇岛市| 株洲市| 文化| 清原| 楚雄市| 会理县| 什邡市| 勐海县| 宜昌市| 高邑县| 漳浦县| 岑巩县| 金坛市| 萍乡市| 浙江省| 法库县| 合山市| 阜城县| 凤凰县| 宽甸| 青浦区| 如东县| 兰西县| 兖州市| 洞口县| 长兴县| 寻甸| 平阴县| 黄龙县| 华坪县| 甘德县| 玛纳斯县| 张北县| 墨竹工卡县| 万荣县| 阿合奇县| 波密县| 南澳县| 孝义市| 那曲县| 油尖旺区|