User Guide

Scaling computation

qianmoQqianmoQ· 更新于 2026-10-09· 阅读 2 分钟· 0 次阅读

登录后可跨设备保存划线和私人笔记登录

Scaling computation

Apache Hamilton enables a variety of tools for allowing you to scale your data processing by integrating with third-party libraries.

Specifically, we have four examples that show how to scale Apache Hamilton both by parallelizing transformations (ray and dask) and running on larger, distributed datasets (pandas on spark, pyspark map UDFs).

  1. Integrating hamilton with pandas on spark.
  2. Integrating hamilton with ray.
  3. Integrating hamilton with dask.
  4. Integrating hamilton with pyspark.

评论

登录后参与评论

正在加载评论…