Starred repositories
DuckDB is an analytical in-process SQL database management system
An agentic skills framework & software development methodology that works.
from vibe coding to agentic engineering - practice makes claude perfect
Videodl: A lightweight video downloader written in pure python. (轻量级视频下载器,优先高清无水印,支持抖音,快手,小红书,B站,TikTok,YouTube,FIFA+,优酷,腾讯,爱奇艺,1905电影网,乐视,芒果,咪咕,PPTV,搜狐,Facebook,Twitter,新浪微博,今日头条,网易公开课,全民K歌,CCTV央视…
SeaTunnel is a distributed, high-performance data integration platform for the synchronization and transformation of massive data (offline & real-time).
Beginner, advanced, expert level Rust training material
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
DataCompare is a tool designed to compare database data. Currently, the databases it supports stably include: PostgreSQL, Oracle, and MySQL; databases under development for support (among the top t…
Pentaho Data Integration ( ETL ) a.k.a Kettle
A fancy, easy-to-use and reactive self-hosted docker compose.yaml stack-oriented manager
etl engine 轻量级 跨平台 流批一体ETL引擎 数据抽取-转换-装载 ETL engine lightweight cross platform batch flow integration ETL engine data extraction transformation loading
Shark Shell 是一个用于管理和自动化常见系统任务的脚本集合。它汇集了在工作中和业余时间开发的各种 Shell 脚本,旨在简化软件安装、配置、启动等任务。该项目涵盖了常见的系统服务与应用的部署,并通过不断的迭代开发,确保脚本的有效性和易用性。
ZHCGitHub / ha_xiaomi_home
Forked from XiaoMi/ha_xiaomi_homeXiaomi Home Integration for Home Assistant
Xiaomi Home Integration for Home Assistant
A modern and practical elasticsearch GUI client | 一个现代、实用的ES本地客户端 💕🎉
A modern and practical kafka GUI client 💕🎉Kafka-King 是一款现代化、实用的 Kafka GUI 客户端,旨在通过直观的桌面界面简化 Apache Kafka 管理。作为一款跨平台应用程序,它为开发人员和管理员提供了强大的工具,可与 Kafka 集群交互,无需依赖命令行界面或基于 Web 的解决方案。
《利用Python进行数据分析·第2版》
Materials and IPython notebooks for "Python for Data Analysis" by Wes McKinney, published by O'Reilly Media
Python - 100天从新手到大师
Flume Source to import data from SQL Databases
ZHCGitHub / Spark_Personas
Forked from Chihuataneo/Spark_PersonasSpark中实现用户画像系统价值度、忠诚度、流失预警、活跃度等模型
Learning Apache spark,including code and data .Most part can run local.
Code for processing AVRO data in Spark Streaming + Kafka (DirectKafka approach with custom offset management in zookeeper)
Kafka stream for Spark with storage of the offsets in ZooKeeper
Code examples that show to integrate Apache Kafka 0.8+ with Apache Storm 0.9+ and Apache Spark Streaming 1.1+, while using Apache Avro as the data serialization format.