Loadingone moment…
Screening then two technical rounds heavy on SQL joins/aggregations and PySpark basics + distributed computing.
Cloud / platform
IBM Cloud + AWS/Azure/GCP per client; Spark, Hive, DataStage, Db2
Reported difficulty
Not reported
Technical screening → SQL (joins/aggregations, window, nested) + PySpark (filtering/joining/top-N, distributed computing) → behavioral.
Medium PySpark interview-question series (worked solutions) reflecting IBM's PySpark emphasis.
Answered questions on what IBM actually asks — each with the short answer, the answer that gets you rejected, and the follow-up.
Then build one and have it reviewed: design a daily orders pipeline or model a marketplace. No account.