-
BV TECH
- Italy
-
06:25
(UTC +02:00) - giovannipasq.github.io
- in/giovanniPasq
- https://scholar.google.com/citations?user=7KPqcUkAAAAJ
Stars
How Python does AI. Agents, realtime voice, image generation, embeddings. Every model, every interface, typed end to end.
Convert PDF to markdown + JSON quickly with high accuracy
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
An extremely fast Python linter and code formatter, written in Rust.
Langflow is a powerful tool for building and deploying AI-powered agents and workflows.
A fast, helpful, and open-source document parser
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while control…
SGLang is a high-performance serving framework for large language models and multimodal models.
DSPy: The framework for programming—not prompting—language models
Polyglot document intelligence with a Rust core: extract text, metadata, images, tables, and structured data from 106 formats across 140 file extensions, plus code intelligence for 371 languages. F…
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
An open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data. Supports NLP, pattern matching, and customizable pipelines.
Python tool for converting files and office documents to Markdown.
🪢 Open source agent evals & observability: Trace, evaluate, and improve LLM applications with one open platform.
Open-source toolkit for reliable RAG pipelines: convert PDFs to Markdown, clean documents, inspect chunks, compare chunking strategies, and enrich metadata for LLM applications.
Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music/song recognition, language detection and timestamp prediction.
An index of the LangChain + LangGraph ecosystem: concepts, projects, tools, templates, and guides for LLM & multi-agent apps.
tiktoken is a fast BPE tokeniser for use with OpenAI's models.
A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.
HyDE: Precise Zero-Shot Dense Retrieval without Relevance Labels
FastAPI framework, high performance, easy to learn, fast to code, ready for production
💥 Fast State-of-the-Art Tokenizers optimized for Research and Production
Chat with your database or your datalake (SQL, CSV, parquet). PandasAI makes data analysis conversational using LLMs and RAG.
🦛 CHONK docs with Chonkie ✨ — The lightweight ingestion library for fast, efficient and robust RAG pipelines
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthr…
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.