Stars
- All languages
- Astro
- Awk
- C
- C#
- C++
- CSS
- CUE
- Clojure
- CoffeeScript
- Dart
- Dockerfile
- Elixir
- Emacs Lisp
- Erlang
- Fluent
- Go
- Groovy
- HCL
- HTML
- Haskell
- Java
- JavaScript
- Jinja
- Jsonnet
- Jupyter Notebook
- Kotlin
- Lua
- MDX
- Markdown
- Mojo
- OCaml
- Objective-C
- Open Policy Agent
- PLpgSQL
- Perl
- Pony
- Pug
- Puppet
- Python
- Raku
- Rich Text Format
- Ruby
- Rust
- SCSS
- Scala
- Shell
- Swift
- Tcl
- TypeScript
- Vim Script
- XSLT
- YAML
- Zig
- jq
A simple screen parsing tool towards pure vision based GUI agent
Time series Timeseries Deep Learning Machine Learning Python Pytorch fastai | State-of-the-art Deep Learning library for Time Series and Sequences in Pytorch / fastai
A better notebook for Scala (and more)
Apache Hamilton helps data scientists and engineers define testable, modular, self-documenting dataflows, that encode lineage/tracing and metadata. Runs and scales everywhere python does.
Logica is a logic programming language that compiles to SQL. It runs on DuckDB, Google BigQuery, PostgreSQL and SQLite.
Python Helper library for Jupyter Notebooks
Includes notes on using Apache Spark, with drill down on Spark for Physics, how to run TPCDS on PySpark, how to create histograms with Spark. Also tools for stress testing, measuring CPUs' performa…
This is a repo documenting the best practices in PySpark.
New Generation Opensource Data Stack Demo
Comprehensive Vector Data Tooling. The universal interface for all vector database, datasets and RAG platforms. Easily export, import, backup, re-embed (using any model) or access your vector data …
Pushdown compute from Snowflake to DuckDB running on your infrastructure
📓 A series of Jupyter notebooks to demonstrate the functionality of Apache Calcite