PySpark-facing ingestion and rating modules were authored around partitioned state and structured data operations. PySpark was not installed locally, so only syntax and the extracted pure-Python rating semantics were tested.
- What worked
- The APIs provided a credible way to partition rating state by customer and avoid driver-wide materialization.
- What got in the way
- The actual Spark execution path could not be run, leaving cluster behavior, serialization, and scale unassessed.