Skip to content
View pawelmhm's full-sized avatar

Organizations

@scrapinghub

Block or report pawelmhm

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
36 stars written in Python
Clear filter

Scrapy, a fast high-level web crawling & scraping framework for Python.

Python 63,440 11,820 Updated Jul 27, 2026

An interactive TLS-capable intercepting HTTP proxy for penetration testers and software developers.

Python 44,469 4,645 Updated Jul 18, 2026

Asynchronous HTTP client/server framework for asyncio and Python

Python 16,506 2,358 Updated Jul 27, 2026

ChatterBot is a machine learning, conversational dialog engine for creating chat bots

Python 14,505 4,423 Updated Jun 19, 2026

Visual scraping for Scrapy

Python 9,505 1,387 Updated Jun 26, 2024

The property-based testing library for Python

Python 8,825 662 Updated Jul 27, 2026

Event-driven networking engine written in Python.

Python 5,974 1,215 Updated Jul 27, 2026

Library for building WebSocket servers and clients in Python

Python 5,703 603 Updated Jul 27, 2026

Command line interface to the freedesktop.org trashcan.

Python 4,444 206 Updated Jul 17, 2026

Lightweight, scriptable browser as a service with an HTTP API

Python 4,190 512 Updated Aug 2, 2024

Up-to-date simple useragent faker with real world database

Python 4,053 538 Updated Mar 29, 2026

High level Python client for Elasticsearch

Python 3,869 790 Updated Apr 18, 2025

WebSocket client for Python

Python 3,708 777 Updated May 4, 2026
Python 3,695 334 Updated Sep 10, 2020

python parser for human readable dates

Python 2,846 507 Updated Jul 23, 2026

WebSocket and WAMP in Python for Twisted and asyncio

Python 2,542 770 Updated Jul 15, 2026

🎭 Playwright integration for Scrapy

Python 1,435 160 Updated Jul 23, 2026

Parsel lets you extract data from XML/HTML documents using XPath or CSS selectors

Python 1,347 165 Updated Jul 27, 2026

The web framework for inventors

Python 1,228 73 Updated Jul 2, 2026

NO LONGER MAINTAINED - A Flask extension for creating simple ReSTful JSON APIs from SQLAlchemy models.

Python 1,004 295 Updated May 16, 2020

Extract embedded metadata from HTML markup

Python 967 121 Updated Apr 1, 2026

HTTP API for Scrapy spiders

Python 882 161 Updated Jun 29, 2026

Zyte Smart Proxy Manager (formerly Crawlera) middleware for Scrapy

Python 363 92 Updated May 4, 2026

Generator of User-Agent header

Python 347 58 Updated Sep 14, 2025

Scrapy spider middleware to ignore requests to pages containing items seen in previous crawls

Python 276 48 Updated Feb 26, 2025

JSON (de)serialization, GraphQL and JSON schema generation using Python typing.

Python 241 17 Updated Jul 13, 2026

Convert Javascript code to an XML document

Python 188 23 Updated Mar 14, 2022

Python bindings to the Brotli compression library

Python 152 31 Updated Mar 5, 2026

Scrapinghub Command Line Client

Python 129 80 Updated Jul 22, 2026
Next