Processing and Analytics
The two pages here cover the serverless end of AWS analytics — querying and transforming data without provisioning a cluster.
Start with Athena and AWS Glue for how data is crawled, catalogued, transformed and then queried in place. Read Amazon Athena vs Amazon Redshift when deciding whether a workload justifies a warehouse.
For cluster-based processing of the same data — Spark, Hive, Flink — see Amazon EMR.