Databricks · databricks.com · checked today
Senior Forward Deployed Engineer (Technical Data Architect)
<p><strong>Req ID: CSQ427R156</strong></p> <p><strong>Location: London, United Kingdom - Hybrid </strong></p> <p><strong>Job Role: Senior Forward Deployed Engineer - Technical Architect level</strong></p>
Skills, with evidence
- Python
Comfortable writing code in either Python, Scala, JavaScript/TypeScript, and modern frameworks
must have - Spark
Deep experience with distributed computing with Apache Spark™ and knowledge of Spark runtime internals
must have - Scala
Comfortable writing code in either Python, Scala, JavaScript/TypeScript, and modern frameworks
must have · not practised here - AWS
Working knowledge of two or more common Cloud ecosystems (AWS, Azure, GCP) with expertise in at least one
must have · not practised here - Azure
Working knowledge of two or more common Cloud ecosystems (AWS, Azure, GCP) with expertise in at least one
must have · not practised here - GCP
Working knowledge of two or more common Cloud ecosystems (AWS, Azure, GCP) with expertise in at least one
must have · not practised here - Governance & security
Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog.
not practised here - Warehousing
Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog.
- Machine learning
<li>Working knowledge of MLOps, ML/AI models and AI APIs</li>
not practised here
Your plan
- Python: the data-wrangling round≈ 4 h
Python
- Data modelling: the round most people fail≈ 4 h
Warehousing
- Spark: read the plan Spark actually ran≈ 2 h
Spark
- applyInPandas per country: what Spark ships to Python, and the native rewriteAdvanced
- Filters you wrote in the wrong place: where Catalyst moves themAdvanced
- snappy, gzip or zstd: measure the Parquet codec trade-off yourselfAdvanced
- A correlated COUNT(*) subquery: the join Spark runs, and the count bug it avoidsAdvanced
- Find the four suspects in a slow pipeline (one of them is innocent)Advanced
- Say it out loud≈ 1 h
15 drills · Advanced + Intermediate≈ 10 hours
Not covered by the plan: Scala, AWS, Azure, GCP, Governance & security, Machine learning.
Readiness
Counted from drills you have completed anywhere on D8LooP.
leaves tomorrowremoved the moment Databricks closes it
