What's new on the cloud for data engineers - part 9 (01-03.2023)

Have you missed any cloud data engineering-related news in the last 3 months? No worries, I got you covered with the new part of the "What's new on the cloud for data engineers..." series.

4-day workshop · In-person or online

What would it take for you to trust your Databricks pipelines in production?

A 3-day bug hunt on a 3-person team costs up to €7,200 in lost engineering time. This workshop teaches you to prevent that — unit tests, data tests, and integration tests for PySpark and Databricks Lakeflow, including Spark Declarative Pipelines.

Unit, data & integration tests
Medallion architecture & Lakeflow SDP
Max 10 participants · production-ready templates
See the full curriculum → €7,000 flat fee · cohort of up to 10
Bartosz Konieczny
Bartosz
Konieczny

This 9th part covers all that happened between 17.12.2022 and 10.03.2023. You'll see that despite the Christmas period, the data services got a lot of exciting updates! As previously, I highlighted the most interesting news.

AWS

Athena

Aurora

Batch

CloudWatch

Database Migration Service

DocumentDB

DynamoDB

ElastiCache

EMR

Serverless:

Others:

Glue

Crawlers:

Streaming:

Studio:

Others:

Kendra

Kinesis

Data Streams:

Firehose:

Lambda

Processing:

Ops/Others:

Managed Workflows for Apache Airflow

Neptune

Redshift

Serverless:

Others:

S3

Storage Lens:

Security:

Other features:

Storage Gateway

MemoryDB

OpenSearch

RDS

Oracle:

SQL Server:

PostgreSQL:

MariaDB:

MySQL + PostgreSQL:

Global:

Snow Family

Step Functions

Timestream

QuickSight

Azure

Backup

Cache for Redis

Cosmos DB

PostgreSQL

Misc:

Data Explorer

Database Migration

Databricks

Event Grid

Functions

Purview

SQL Database

Hyperscale:

PostgreSQL:

SQL Managed Instance:

MySQL:

Security:

Misc:

Storage Account

Stream Analytics

GCP

BigQuery

Administration/OPS:

IO:

SQL:

Security:

Omni:

Console:

Other features:

BigQuery Transfer Service

BigLake

Cloud Composer

Bug fixes:

Cloud Functions

Cloud SQL

SQL Server:

MySQL:

PostgreSQL:

PostgreSQL and MySQL:

Global:

Cloud Storage

Data Fusion

Data Loss Protection

New detectors and connections:

Others:

Dataflow

Dataplex

Dataproc

Datastream

Firestore

IAM

Spanner

Storage Transfer Service

This time I noticed less targeted-changes. In the previous updates, I had a feeling that the cloud providers were working on a particular topic each time, such as table file formats or streaming. Here, the changes are different but definitively, the new autoscaling on BigQuery, DocumentDB stream, or cross-cloud connectivity, are interesting features. What's yours?

Data Engineering Design Patterns

Looking for a book that defines and solves most common data engineering problems? I wrote one on that topic! You can read it online on the O'Reilly platform, or get a print copy on Amazon.

I also help solve your data engineering problems contact@waitingforcode.com đź“©