49 repos across 3 sub-areas
Distributed computing frameworks and data pipeline tools built primarily in Scala and Java. The cluster centers on Apache Spark for large-scale data processing, Akka for actor-based concurrency and reactive systems, and Akka HTTP for building reactive web services. These projects enable building scalable data workflows, streaming applications, and real-time processing systems, with supporting libraries like Alpakka for reactive integrations and Linkis for data engine orchestration.
Cluster 462200
27 repos
Cluster 462201
20 repos
Scala and JVM Streaming Systems
2 repos
Libraries, tools, and frameworks for building distributed streaming and data processing systems on the JVM, with a heavy emphasis on Scala. This cluster includes Akka-based reactive streaming infrastructure, Kafka integration patterns, code generation and formatting tooling, and documentation systems—all centered on the Scala and Java ecosystem. Developers exploring this area will find configuration management, build tool plugins, code samples, and supporting utilities for constructing scalable, event-driven applications.