랭킹으로 돌아가기

hhursev/recipe-scrapers

Pythondocs.recipe-scrapers.com

Python package for scraping recipes data

data-extractionfoodhtml-parsingjson-ldlibrarymealmicrodataopengraphpythonrecipe-schemarecipe-scraperschema-org
스타 성장
스타
2.2k
포크
664
주간 성장
이슈
108
1k2k
2023년 1월2024년 3월2025년 5월2026년 7월
아티팩트PyPIpip install recipe-scrapers
관련 저장소
firecrawl/firecrawl

The API to search, scrape, and interact with the web at scale. 🔥

TypeScriptnpmGNU Affero General Public License v3.0aicrawler
firecrawl.dev
154.1k8.8k
D4Vinci/Scrapling

🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!

PythonPyPIBSD 3-Clause "New" or "Revised" Licensecrawlercrawling
scrapling.readthedocs.io/en/latest/
70.6k7k
ScrapeGraphAI/Scrapegraph-ai

Python scraper based on AI

PythonPyPIMIT Licensescrapingscraping-python
scrapegraphai.com
28.5k2.8k
getmaxun/maxun

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥

TypeScriptnpmGNU Affero General Public License v3.0automationno-code
maxun.dev
16.7k1.4k
MontFerret/ferret

Declarative data automation language and Go runtime for structured extraction workflows.

GoGo ModulesApache License 2.0golangquery-language
ferretlang.org
6k326
vi3k6i5/flashtext

Extract Keywords from sentence or Replace keywords in sentences.

PythonPyPIMIT Licensesearch-in-textkeyword-extraction
5.7k596
browser-act/skills

Browser automation CLI built for AI agents. Break through anti-bot walls, hand off to humans across platforms when stuck. Parallel multi-task execution, independent multi-session operation, isolated multi-account browsing.

PythonPyPIMIT Licenseai-agentsautomation
browseract.com
4.6k217
brightdata/brightdata-mcp

A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access.

JavaScriptnpmMIT Licensellmmcp
brightdata.com
2.5k316
saifyxpro/HeadlessX

The undetected self-hosted browser automation platform. Powered by Camoufox (Firefox) for 0% detection rates. Built for speed, privacy, and scalability.

TypeScriptnpmMIT Licenseautomationbrowser-automation
headlessx.saify.me
2k255
shcherbak-ai/contextgem

ContextGem: Effortless LLM extraction from documents

PythonPyPIApache License 2.0aicontract-analysis
contextgem.dev
1.9k158
0xMassi/webclaw

Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust. CLI, REST API, and MCP server.

Rustcrates.ioGNU Affero General Public License v3.0ai-agentscli
webclaw.io
1.8k194
JonathanLink/PDFLayoutTextStripper

Converts a pdf file into a text file while keeping the layout of the original pdf. Useful to extract the content from a table in a pdf file for instance. This is a subclass of PDFTextStripper class (from the Apache PDFBox library).

JavaMavenApache License 2.0layouttext
jonathanlink.ch/PDFLayoutTextStripper.html
1.6k213