6-day longest streak
-
scrapy ★ PINNED ⑂
Scrapy, a fast high-level web crawling & scraping framework for Python.
Python ★ 1 2d agoExplain → -
python-community-map ⑂
A map full of lovely Python communities ❤️🐍🌎
HTML ★ 3 7y agoExplain → -
frontera ⑂
A scalable frontier for web crawlers
★ 2 1y agoExplain → -
extruct ⑂
Extract embedded metadata from HTML markup
★ 1 1y agoExplain → -
tabula-py ⑂
Simple wrapper of tabula-java: extract table from PDF into pandas DataFrame
Python ★ 1 7y agoExplain → -
parsel ⑂
Parsel lets you extract data from XML/HTML documents using XPath or CSS selectors
Python ★ 1 2y agoExplain → -
aiohttp ⑂
Async http client/server framework (asyncio)
Python ★ 1 9y agoExplain → -
TwitchIO ⑂
TwitchIO - An Async Bot/API wrapper for Twitch made in Python.
★ 0 6y agoExplain → -
Cyberpunk-Modding-Docs ⑂
Modding tutorials and documentation for Cyberpunk 2077
★ 0 7mo agoExplain → -
godot ⑂
Godot Engine – Multi-platform 2D and 3D game engine
★ 0 7mo agoExplain → -
vscode-mre
No description.
JavaScript ★ 0 11mo agoExplain → -
scrapy-project
No description.
Python ★ 0 11mo agoExplain → -
scrapyrt ⑂
HTTP API for Scrapy spiders
★ 0 1y agoExplain → -
vscode ⑂
Visual Studio Code
TypeScript ★ 0 1y agoExplain → -
flake8-scrapy ⑂
A Flake8 plugin to catch common issues on Scrapy spiders
★ 0 1y agoExplain → -
ruff ⑂
An extremely fast Python linter and code formatter, written in Rust.
★ 0 1y agoExplain → -
paramiko-sftp-example
Paramiko-based minimal SFTP client and server examples
Python ★ 0 1y agoExplain → -
duplicate-url-discarder ⑂
No description.
★ 0 1y agoExplain → -
web-scraping-tutorial-project ⑂
https://docs.zyte.com/web-scraping/tutorial/index.html
Python ★ 0 1y agoExplain → -
python-zyte-api-77
No description.
Python ★ 0 1y agoExplain → -
extract-summit-contest-solutions ⑂
Example solutions for the practice and contest websites of the code contest of Web Data Extraction Summit.
Python ★ 0 1y agoExplain → -
form2request ⑂
AI-powered Python 3.8+ library to build HTTP requests out of HTML forms.
★ 0 1y agoExplain → -
sklearn-crfsuite ⑂
scikit-learn inspired API for CRFsuite
Python ★ 0 2y agoExplain → -
Formasaurus ⑂
Formasaurus tells you the type of an HTML form and its fields using machine learning
HTML ★ 0 1y agoExplain → -
pycones24 ⑂
PyConEs 2024 Vigo
★ 0 2y agoExplain → -
OpenNutriTracker ⑂
🍴 OpenNutriTracker is a free and open source calorie tracker with a focus on simplicity and privacy.
★ 0 2y agoExplain → -
scrapy-feedexporter-onedrive ⑂
Export to OneDrive
★ 0 2y agoExplain → -
transitous ⑂
Free and open public transport routing.
★ 0 2y agoExplain → -
scrapinghub-stack-scrapy ⑂
Software stack with latest Scrapy and updated deps
★ 0 2y agoExplain → -
scrapy-frontera ⑂
More flexible and featured Frontera scheduler for Scrapy
★ 0 1y agoExplain → -
andi ⑂
Library for annotation-based dependency injection
★ 0 1y agoExplain → -
zyte-spider-templates-project ⑂
No description.
★ 0 1y agoExplain → -
aiohttp-swagger ⑂
Swagger API Documentation builder for aiohttp server
Python ★ 0 9y agoExplain → -
zyte-spider-templates ⑂
Spider templates for automatic crawlers.
★ 0 1y agoExplain → -
url-matcher ⑂
No description.
Python ★ 0 2y agoExplain → -
scrapy-spider-metadata ⑂
No description.
★ 0 1y agoExplain → -
scrapy-xlsx ⑂
XLSX exporter for Scrapy
★ 0 3y agoExplain → -
jsonlines ⑂
python library to simplify working with jsonlines and ndjson data
★ 0 2y agoExplain → -
languagetool ⑂
Style and Grammar Checker for 25+ Languages
★ 0 3y agoExplain → -
hatch-msgfmt ⑂
A hatch msgfmt plugin, integrating msgfmt.py to convert .po files to .mo during build.
★ 0 3y agoExplain → -
reuse-tool ⑂
reuse is a tool for compliance with the REUSE recommendations.
★ 0 3y agoExplain → -
sddm ⑂
QML based X11 and Wayland display manager
C++ ★ 0 3y agoExplain → -
scrapy-splash ⑂
Scrapy+Splash for JavaScript integration
Python ★ 0 1y agoExplain → -
weblate ⑂
Web based localization tool with tight version control integration.
★ 0 3y agoExplain → -
scrapy-zyte-api ⑂
Zyte Data API integration for Scrapy
★ 0 1y agoExplain → -
xtractmime ⑂
https://mimesniff.spec.whatwg.org/ implementation for Python
★ 0 2y agoExplain → -
scrapy-poet ⑂
Page Object pattern for Scrapy
★ 0 1y agoExplain → -
web-poet ⑂
Web scraping Page Objects core library
★ 0 11mo agoExplain → -
python-zyte-api ⑂
Python client for Zyte Data API
★ 0 11mo agoExplain → -
pytest ⑂
The pytest framework makes it easy to write small tests, yet scales to support complex functional testing
★ 0 4y agoExplain → -
zyte-common-items ⑂
Contains the common item definitions used in Zyte.
★ 0 11mo agoExplain → -
scrapy-feedstock ⑂
A conda-smithy repository for scrapy.
★ 0 2y agoExplain → -
python-scrapinghub ⑂
A client interface for Scrapinghub's API
Python ★ 0 1y agoExplain → -
dicttoxml ⑂
Simple library to convert a Python dictionary or other native data type into a valid XML string.
★ 0 6y agoExplain → -
curlconverter ⑂
convert curl commands to python, javascript, php, R, Go, Rust, Dart, JSON, Ansible
★ 0 6y agoExplain → -
scrapy.org ⑂
The scrapy.org website
HTML ★ 0 2y agoExplain → -
website ⑂
// foss.events: The most comprehensive collection of FOSS events in Europe
★ 0 5y agoExplain → -
zyte-smartproxy-headless-proxy ⑂
A complimentary proxy to help to use Crawlera with headless browsers
★ 0 5y agoExplain → -
zyte-smartproxy-clients ⑂
Crawlera HTTPS clients collection
★ 0 5y agoExplain → -
gsoc-proposal ⑂
GSoC proposal for scrapy
★ 0 5y agoExplain → -
shublang ⑂
Pluggable DSL that uses pipes to perform a series of linear transformations to extract data
★ 0 5y agoExplain → -
zyte-autoextract ⑂
Python clients for Scrapinghub AutoExtract API
★ 0 5y agoExplain → -
spidyquotes ⑂
Example site for web scraping tutorials
★ 0 5y agoExplain → -
scrapinghub-entrypoint-scrapy ⑂
Scrapy entrypoint for Scrapinghub job runner
★ 0 1y agoExplain → -
scrapy-rotating-proxies ⑂
use multiple proxies with Scrapy
★ 0 5y agoExplain → -
autopager ⑂
Detect and classify pagination links
★ 0 5y agoExplain → -
scrapy-jsonschema ⑂
Scrapy schema validation pipeline and Item builder using JSON Schema
★ 0 5y agoExplain → -
scrapely ⑂
A pure-python HTML screen-scraping library
★ 0 6y agoExplain → -
quotesbot ⑂
This is a sample Scrapy project for educational purposes
★ 0 6y agoExplain → -
loginform ⑂
Fill HTML login forms automatically
★ 0 5y agoExplain → -
Spider-Sense ⑂
A browser extension to monitor your spiders deployed on Scrapy Cloud.
★ 0 5y agoExplain → -
twisted_hang ⑂
Hack day project. Figure out if the main thread is hanging, and if so, what's causing it to hang.
★ 0 13y agoExplain → -
scrapy-bench ⑂
A CLI for benchmarking Scrapy.
★ 0 5y agoExplain → -
itemadapter ⑂
Common interface for data container classes
★ 0 11mo agoExplain → -
itemloaders-feedstock ⑂
A conda-smithy repository for itemloaders.
★ 0 6y agoExplain → -
itemloaders ⑂
Library to populate items using XPath and CSS with a convenient API
★ 0 2y agoExplain → -
splash ⑂
Lightweight, scriptable browser as a service with an HTTP API
Python ★ 0 5y agoExplain → -
steam-cli ⑂
Command-line interface to install and execute Steam games
Python ★ 0 6y agoExplain → -
pycrypto ⑂
The Python Cryptography Toolkit
Python ★ 0 7y agoExplain → -
pyopenssl ⑂
A Python wrapper around the OpenSSL library
★ 0 6y agoExplain → -
staged-recipes ⑂
A place to submit conda recipes before they become fully fledged conda-forge feedstocks
★ 0 5y agoExplain → -
python-gsoc.github.io ⑂
Website and ideas page for Python's Google Summer of Code efforts
★ 0 6y agoExplain → -
protego ⑂
A pure-Python robots.txt parser with support for modern conventions.
DIGITAL Command Language ★ 0 1y agoExplain → -
shub-workflow ⑂
No description.
★ 0 6y agoExplain → -
queuelib ⑂
Collection of persistent (disk-based) queues
★ 0 5y agoExplain → -
scrapy-inline-requests ⑂
A decorator to write coroutine-like spider callbacks.
Python ★ 0 10y agoExplain → -
twisted ⑂
Event-driven networking engine written in Python.
Python ★ 0 2y agoExplain → -
arche ⑂
Analyze scraped data
Python ★ 0 7y agoExplain → -
spider-feeder ⑂
No description.
Python ★ 0 5y agoExplain → -
shub ⑂
Scrapinghub Command Line Client
Python ★ 0 2y agoExplain → -
eli5 ⑂
A library for debugging/inspecting machine learning classifiers and explaining their predictions
Jupyter Notebook ★ 0 7y agoExplain → -
sphinx ⑂
Main repository for the Sphinx documentation builder
Python ★ 0 7y agoExplain → -
price-parser ⑂
Extract price amount and currency symbol from a raw text string
Python ★ 0 5y agoExplain → -
html-text ⑂
Extract text from HTML
HTML ★ 0 6y agoExplain → -
w3lib ⑂
Python library of web-related functions
Python ★ 0 2y agoExplain → -
charlas ⑂
Repositorio de datos y slides de las charlas llevadas a cabo por el grupo de Python Vigo
CSS ★ 0 7y agoExplain → -
wgrep ⑂
Web grep: search all rendered resources used by a URI
JavaScript ★ 0 7y agoExplain → -
js2xml ⑂
Convert Javascript code to an XML document
Python ★ 0 4y agoExplain → -
cssselect ⑂
CSS Selectors for Python
Python ★ 0 2y agoExplain → -
spidermon ⑂
Monitor Scrapy Cloud spiders
Python ★ 0 5y agoExplain → -
dateparser ⑂
python parser for human readable dates
Python ★ 0 1y agoExplain → -
scrapy-crawlera ⑂
Crawlera middleware for Scrapy
Python ★ 0 1y agoExplain → -
boltons ⑂
🔩 Like builtins, but boltons. 220+ constructs, recipes, and snippets extending (and relying on nothing but) the Python standard library. Nothing like Michael Bolton.
Python ★ 0 7y agoExplain → -
pycryptodome ⑂
A self-contained cryptographic library for Python
Python ★ 0 7y agoExplain → -
validators ⑂
Python Data Validation for Humans™.
Python ★ 0 7y agoExplain → -
hunspell-gl ⑂
hunspell-gl
Python ★ 0 7y agoExplain → -
libnumbertext ⑂
Number to number name and money text conversion libraries in C++, Java, JavaScript and Python & LibreOffice Calc Extension
M4 ★ 0 8y agoExplain → -
phpstan-laravel ⑂
Laravel plugins for PHPStan
PHP ★ 0 8y agoExplain → -
translate ⑂
Useful localization tools with Python API for building localization & translation systems
Python ★ 0 8y agoExplain → -
l10n-guide ⑂
Localisation guide
★ 0 8y agoExplain → -
ProxyBroker ⑂
Proxy [Finder | Checker | Server]. HTTP(S) & SOCKS
Python ★ 0 9y agoExplain →
No repos match these filters.