AWS msk: Restructured MSK Data Delivery documentation with feature overview
Summary
Replaced IAM permissions content with detailed descriptions of Amazon MSK Data Delivery capabilities for Iceberg tables and S3 buckets.
Security assessment
The change replaces IAM permission documentation with feature descriptions of data delivery capabilities. No security vulnerabilities, configurations, or best practices are addressed or added.
Evidence
+# Amazon MSK Data Delivery
Diff
diff --git a/msk/latest/developerguide/msk-data-delivery-iam.md b/msk/latest/developerguide/msk-data-delivery-iam.md index 8c028fbde..61b71a0b5 100644 --- a//msk/latest/developerguide/msk-data-delivery-iam.md +++ b//msk/latest/developerguide/msk-data-delivery-iam.md @@ -1 +1 @@ -[View a markdown version of this page](msk-data-delivery-iam.md) +[View a markdown version of this page](msk-data-delivery.md) @@ -3 +3 @@ -[](/pdfs/msk/latest/developerguide/MSKDevGuide.pdf#msk-data-delivery-iam "Open PDF") +[](/pdfs/msk/latest/developerguide/MSKDevGuide.pdf#msk-data-delivery "Open PDF") @@ -7 +7,11 @@ -# IAM permissions for Channel +# Amazon MSK Data Delivery + +With Amazon MSK data delivery, you can deliver Apache Kafka data from Amazon MSK Express brokers directly to Amazon S3, without connectors or additional infrastructure to manage. Amazon MSK Express automatically handles scaling, retries, and backpressure, and manages routine operations such as capacity scaling and version upgrades without introducing delivery gaps. Because these are native broker capabilities, they add no broker egress throughput, so you avoid the incremental infrastructure costs that scaling connector-based pipelines typically incurs and match capacity to actual workload demand rather than provisioning for peak. Each capability supports throughput of up to 10 GBps. + +The two capabilities are: + + * **Data delivery to streaming tables for Apache Iceberg** — With Amazon MSK Data Delivery, you can continuously materialize Apache Kafka topics as Apache Iceberg tables on Amazon S3 Tables. Intelligent inline compaction eliminates the performance impact of small files and keeps query performance predictable without sacrificing data freshness. Built-in coordination resolves concurrent writer conflicts across high-throughput consumers. Amazon S3 Tables automatically handles ongoing table maintenance, including compaction, snapshot expiration, and unreferenced file cleanup. + + * **Data delivery to Amazon S3 general purpose buckets** — With Amazon MSK Data Delivery, you can deliver Apache Kafka data in the source format to Amazon S3 general purpose buckets for downstream processing, with end-to-end reliability for mission-critical workloads. Use it to land Kafka data in Amazon S3 for use cases such as log archival, compliance retention, Kafka replay, and training AI/ML models. This approach removes the need to build self-managed connector pipelines that grow costly and operationally complex as workloads scale. + + @@ -9 +18,0 @@ -There are two distinct sets of permissions: the **Channel lifecycle management permissions** that you need to create and manage Channels, and the **service execution role** that the Channel assumes to deliver data. @@ -13 +22 @@ There are two distinct sets of permissions: the **Channel lifecycle management p - * [Channel lifecycle management permissions](./msk-data-delivery-iam-lifecycle.html) + * [data delivery for streaming tables to Apache Iceberg](./msk-data-delivery-iceberg.html) @@ -15 +24 @@ There are two distinct sets of permissions: the **Channel lifecycle management p - * [Service execution role](./msk-data-delivery-iam-service-role.html) + * [data delivery to Amazon S3 general purpose buckets](./msk-data-delivery-s3.html) @@ -26 +35 @@ To use the Amazon Web Services Documentation, Javascript must be enabled. Please -Create your first Channel +Troubleshooting @@ -28 +37 @@ Create your first Channel -Channel lifecycle management permissions +data delivery for streaming tables to Apache Iceberg