mozilla-services/heka

Gohekad.readthedocs.org
macOSLinux

alternativa a SplunkElastic Stack

DEPRECATED: Data collection and processing made easy.

data-collectionstream-processinglogginggodeprecated
Crescimento de estrelas
Estrelas
3.4k
Forks
514
Crescimento semanal
+0
Issues
219
1k2k3k
jan. de 2023mar. de 2024jun. de 2025set. de 2026
ArtefatosGo Modules
README

This project is deprecated. Please see this email for more details.

Heka

Data Acquisition and Processing Made Easy

Heka is a tool for collecting and collating data from a number of different sources, performing "in-flight" processing of collected data, and delivering the results to any number of destinations for further analysis.

Heka is written in Go, but Heka plugins can be written in either Go or Lua. The easiest way to compile Heka is by sourcing (see below) the build script in the root directory of the project, which will set up a Go environment, verify the prerequisites, and install all required dependencies. The build process also provides a mechanism for easily integrating external plug-in packages into the generated hekad. For more details and additional installation options see Installing.

WARNING: YOU MUST SOURCE THE BUILD SCRIPT (i.e. source build.sh) TO BUILD HEKA. Setting up the Go build environment requires changes to the shell environment, if you simply execute the script (i.e. ./build.sh) these changes will not be made.

Resources:

Repositórios relacionados
NaiboWang/EasySpider

A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。

JavaScriptnpmappGNU Affero General Public License v3.0code-freecrawler
easyspider.net
44.5k5.4k
airbytehq/airbyte

Open-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses, lakes, and AI applications. Both self-hosted and Cloud.

PythonPyPIOtherdatapipeline
airbyte.com
22k5.3k
cporter202/API-mega-list

This GitHub repo is a powerhouse collection of APIs you can start using immediately to build everything from simple automations to full-scale applications. One of the most valuable API lists on GitHub—period. 💪

JavaScriptnpmawesomeapiapi-list
7.6k1.5k
firecrawl/firecrawl-mcp-server

🔥 Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.

JavaScriptnpmlibraryMIT Licensebatch-processingclaude
firecrawl.dev
7.4k880
snowplow/snowplow

The leader in Customer Data Infrastructure

ScalaApache License 2.0snowplowanalytics
snowplow.io
7k1.2k
cloudquery/cloudquery

Data pipelines for cloud config and security data. Build cloud asset inventory, CSPM, FinOps, and vulnerability management solutions. Extract from AWS, Azure, GCP, and 70+ cloud and SaaS sources.

GoGo ModulesMozilla Public License 2.0awsgcp
cloudquery.io
6.5k557
jitsucom/jitsu

Jitsu is an open-source Segment alternative. Fully-scriptable data ingestion engine for modern data teams. Set-up a real-time data pipeline in minutes, not days

TypeScriptnpmMIT Licensedata-integrationclickhouse
jitsu.com
5.1k395
vladkens/twscrape

Python library and CLI for X/Twitter scraping with multi-account rotation and built-in rate-limit handling.

PythonPyPIlibraryMIT Licensetwitterscraper
pypi.org/project/twscrape/
2.7k318
brightdata/brightdata-mcp

A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access.

JavaScriptnpmMIT Licensellmmcp
brightdata.com
2.6k325
Altimis/Scweet

Scrape tweets, profiles, followers and following from Twitter/X, no API key needed. Python library with smart multi-account pooling, proxy support and async.

PythonPyPIlibraryMIT Licensescrapertwitter
apify.com/altimis/scweet
1.6k281
pyper-dev/pyper

Concurrent Python made simple

PythonPyPIlibraryMIT Licenseasyncioconcurrency
1.5k32
Decodo/Decodo

HTTP(S)/SOCKS5 rotating residential proxies - code examples & general information.

JavaMavenMIT Licenseproxyproxies
decodo.com
1.2k57