<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Backpressured</title><link>https://backpressured.dev/en/</link><description>Recent content on Backpressured</description><generator>Hugo</generator><language>en</language><lastBuildDate>Thu, 11 Jun 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://backpressured.dev/en/index.xml" rel="self" type="application/rss+xml"/><item><title>The Hidden Cost of CDC Pipelines: How Small Files Create an S3 Request Bomb</title><link>https://backpressured.dev/en/posts/cdc-small-file-hidden-cost/</link><pubDate>Thu, 11 Jun 2026 00:00:00 +0000</pubDate><guid>https://backpressured.dev/en/posts/cdc-small-file-hidden-cost/</guid><description>Athena cost isn&amp;#39;t just about data scanned. Learn how millions of small files from CDC pipelines silently drain your S3 request budget and degrade query performance — and how we fixed it.</description></item><item><title>About</title><link>https://backpressured.dev/en/about/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://backpressured.dev/en/about/</guid><description>&lt;p>Hi, I&amp;rsquo;m Jaehyuk.&lt;/p>
&lt;p>I build and operate data platforms at a fintech company in India. I mostly work with Spark, Flink, Debezium, Airflow, and Iceberg. These days I&amp;rsquo;m interested in modernizing data stacks and building data platforms that are AI-ready.&lt;/p>
&lt;p>This is where I write about things I learn along the way.&lt;/p>
&lt;h2 id="tech-stack">Tech Stack&lt;/h2>
&lt;ul>
&lt;li>&lt;strong>Processing:&lt;/strong> Spark, Flink, dbt&lt;/li>
&lt;li>&lt;strong>Streaming &amp;amp; CDC:&lt;/strong> Debezium, Kafka, Kinesis, NiFi&lt;/li>
&lt;li>&lt;strong>Storage &amp;amp; DW:&lt;/strong> S3, Athena, Iceberg, DynamoDB&lt;/li>
&lt;li>&lt;strong>Orchestration:&lt;/strong> Airflow&lt;/li>
&lt;li>&lt;strong>DevOps &amp;amp; Infra:&lt;/strong> Docker, Jenkins, SAM, CloudFormation, Lake Formation&lt;/li>
&lt;li>&lt;strong>Monitoring:&lt;/strong> Prometheus, Grafana, CloudWatch, Kibana&lt;/li>
&lt;li>&lt;strong>Cloud:&lt;/strong> AWS&lt;/li>
&lt;li>&lt;strong>Languages:&lt;/strong> Python, SQL&lt;/li>
&lt;/ul>
&lt;h2 id="writing">Writing&lt;/h2>
&lt;ul>
&lt;li>&lt;a href="https://blog.afinit.com/cdc-incremental-replication">How CDC Changes Data Platforms: CDC-based Incremental Replication&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://blog.afinit.com/cdc-pipeline-debezium-flink">Why We Redesigned Our CDC Pipeline with Debezium and Flink&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://blog.afinit.com/nifi-apache-flink-sms-pipeline">From NiFi to Apache Flink: Improving a Real-time SMS Pipeline&lt;/a>&lt;/li>
&lt;/ul>
&lt;h2 id="contact">Contact&lt;/h2>
&lt;ul>
&lt;li>&lt;a href="https://linkedin.com/in/jaehyuk-jang-bab185178">LinkedIn&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://github.com/jaehyukjang">GitHub&lt;/a>&lt;/li>
&lt;li>Email: &lt;a href="mailto:jjhyuk92@gmail.com">jjhyuk92@gmail.com&lt;/a>&lt;/li>
&lt;/ul></description></item></channel></rss>