Doikayt Tech Blog
Attempted Bloggery in the 1st degree:
The new stuff !
- 2026-07-27 — Bitwarden, Autofill, and the Tedium of Submitting to Community Calendars (Work in Progress)
- 2026-07-26 — Dummy Blog Three
- 2026-07-25 — Dummy Blog Two
- 2026-07-24 — Dummy Blog One
DataLackey Labs Blog Archive
Legacy posts
-- to show potential consulting clients we've done a thing or two with data pipelines !
Archived posts from the old datalackey.com blog.
- 2020-09-22 — Unit & Integration Testing Kafka and Spark
- 2019-09-05 — Time Travails With Java, Scala and Apache Spark
- 2019-08-24 — Getting Spark 2.4.3 Multi-node (Stand-alone) Cluster Working With Docker
- 2019-08-23 — Spark Structured Streaming Joins With No Watermarks Can Blow Out Your Memory
- 2019-07-01 — Sliding Window Processing: Spark Structured Streaming vs. DStreams
- 2019-06-21 — Exploring Event Time and Processing Time in Spark Structured Streaming
- 2019-06-03 — Spark Windowing and Aggregation Functions for DataFrames
- 2019-04-22 — Can Adding Partitions Improve The Performance of Your Spark Job On Skewed Data Sets?
- 2019-04-16 — Assessing Significant Difference In Pairwise Combinations via the Marascuilo Procedure (in R)
- 2019-04-09 — How To Inspect Attribute Info of Nodes in a JQuery Select List
- 2019-04-08 — Configuring the Xerces XML Parser With Content Model Defaults
- 2019-04-03 — Reducing Integration Hassles With JSON Schema Contracts