Which technology enables distributed processing of large datasets across clusters?
- MySQL
- Apache Spark
- MongoDB
- Redis
Answer: Apache Spark
Apache Spark provides in-memory distributed computing for big data: batch, streaming, ML, graph processing. Faster than Hadoop MapReduce due to in-memory processing. Supports Scala, Python, SQL. Critical for big data engineering and analytics questions.