Replies: 4 comments
From PR #338Section 7 — DiscussionThe |
From issue #335 / PR #336Section 7 — DiscussionThe transient IMDS failure observed in batch 18/18 of a 3.1h run (issue #335) is a concrete example of the cloud infrastructure variability discussed in Section 4.5. The retry mechanism is a reliability mitigation, not a measurement concern — retries happen at container creation time, outside the |
From issue #254§7 — Threats to validity: missing default-strategy baselineThe Sedona default join strategy (no explicit broadcast or repartition hint) was dropped from the experiment suite because iterations consistently timed out, failed, or both — making data collection infeasible within budget and time constraints. This is a threat to external validity: the default strategy is what a user would get without tuning, so its absence means the results cannot speak to out-of-the-box Sedona performance on national-scale joins. The thesis should note that:
|
From PR #343Chapter 7 is currently a stub with discussion goals. The following points are relevant when writing the discussion. Cost model accuracy and RQ2 implicationsPR #343 corrected four cost model gaps that systematically underestimated Databricks/Sedona costs relative to single-node configurations. When discussing RQ2 (scaling against single-node), acknowledge:
|
Uh oh!
There was an error while loading. Please reload this page.
Purpose
Working notes for Chapter 7: Discussion of the master thesis
"Benchmarking Cloud-Native and Traditional Geospatial Technologies."
This chapter is currently a placeholder with stated goals.
Discussion goals (from thesis)
Expected content
How to use this thread
Post notes, draft paragraphs, open questions, and revision requests as comments.
All reactions