- SHANG HAI
Stars
Apache Celeborn is an elastic and high-performance service for shuffle and spilled data.
This project provides example FeatHub (https://github.com/alibaba/feathub) programs
A collection of algorithms for mining data streams
FeatHub - A stream-batch unified feature store for real-time machine learning
Deep Learning on Flink aims to integrate Flink and deep learning frameworks (e.g. TensorFlow, PyTorch, etc) to enable distributed deep learning training and inference on a Flink cluster.
Demos for Flink connectors on Ververica Platform (VVP)
SeaTunnel is a multimodal, high-performance, distributed, massive data integration tool.
Apache Paimon is a lake format that enables building a Realtime Lakehouse Architecture with Flink and Spark for both streaming and batch operations.
dbt enables data analysts and engineers to transform their data using the same practices that software engineers use to build applications.
Apache Flink Kubernetes Operator
flink learning blog. http://www.54tianzhisheng.cn/ 含 Flink 入门、概念、原理、实战、性能调优、源码解析等内容。涉及 Flink Connector、Metrics、Library、DataStream API、Table API & SQL 等内容的学习案例,还有 Flink 落地应用的大型项目案例(PVUV、日志存储、百亿数据实时去…
Clink is a library that provides APIs and infrastructure to facilitate the development of parallelizable feature engineering operators that can be used in both C++ and Java runtime.
AI Flow is an open source framework that bridges big data and artificial intelligence.
Debug your GitHub Actions via SSH by using tmate to get access to the runner system itself.
Alibaba Java Diagnostic Tool Arthas/Alibaba Java诊断利器Arthas
Matrix Multiplication on GPU using Shared Memory considering Coalescing and Bank Conflicts
Apache Beam is a unified programming model for Batch and Streaming data processing.
📚 Freely available programming books
Flink CDC is a streaming data integration tool
BigDL: Distributed TensorFlow, Keras and PyTorch on Apache Spark/Flink & Ray