data-pipeline
21 个项目 · ⭐ 95.0kEmpowering Data Intelligence with Distributed SQL for Sharding, Scalability, and Security Across All Databases.
Change data capture for a variety of databases. Please log issues at https://github.com/debezium/dbz/issues.
High-performance AI pipeline engine with a C++ core and 50+ Python-extensible nodes. Build, debug, and scale LLM workflows with 13+ model providers, 8+ vector databases, and agent orchestration, all from your IDE. Includes VS Code extension, TypeScript/Python SDKs, and Docker deployment.
Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷
Where data access meets operational intelligence
Self-hostable workflow orchestrator for teams whose main work isn't orchestration. Declarative YAML over your scripts, SSH commands, containers, etc; keep workflows separate from business logic. One binary, no database, runs on limited H/W resources. Alternative to Airflow / Cron / Job Scheduler.
Open-source inference server and production cluster for all the models your agent needs.
🔥🔥🔥 Open source Reverse ETL - alternative to hightouch and census.
DataMate is an enterprise-level data processing platform designed for model fine-tuning and RAG retrieval.
Use LLMs to robustly extract web data
共 21 条 · 第 1 / 2 页