Data-integratie
240 projecten in Data & analyse
Toont top 200 van 240 op Atlas Score; verfijn met de filters hierboven.
ShardingSphere
@apacheEmpowering Data Intelligence with Distributed SQL for Sharding, Scalability, and Security Across All Databases.
Kestra
@kestra-ioEvent-driven, language-agnostic platform to create, schedule, and monitor workflows. In code. Coordinate data pipelines and tasks such as ETL and ELT.
Apache Airflow
@apachePlatform to programmatically author, schedule, and monitor workflows.
Vector
@vectordotdevA high-performance observability data pipeline.
multiwoven
@Multiwoven🔥🔥🔥 Open source Reverse ETL - alternative to hightouch and census.
Prefect
@PrefectHQPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
whodb
@clideyWhere data access meets operational intelligence
cloudquery
@cloudqueryData pipelines for cloud config and security data. Build cloud asset inventory, CSPM, FinOps, and vulnerability management solutions. Extract from AWS, Azure, GCP, and 70+ cloud and SaaS sources.
RudderStack
@rudderlabsCollect, unify, transform, and store your customer data, and route it to a wide range of common, popular marketing, sales, and product tools (alternative to Segment).
mage-ai
@mage-ai🧙 Build, run, and manage data pipelines for integrating and transforming data.
sqlmesh
@SQLMeshScalable and efficient data transformation framework - backwards compatible with dbt.
taipy
@AvaigaTurns Data and AI algorithms into production-ready web applications in no time.
duckle
@slothflowlabsOpen-source ETL/ELT you deploy on your own servers or cloud. Built on DuckDB: no-code/low-code visual pipelines or SQL, 385 components, dbt, CDC, data quality, reverse ETL, lineage, MCP for AI agents. No vendor cloud, no per-row billing.
flink-cdc
@apacheFlink CDC is a streaming data integration tool
myduckserver
@apecloudUnified MySQL, Postgres & FlightSQL Server, Powered by DuckDB.
reductstore
@reductstoreHigh-performance, time-indexed object storage for robotics and industrial IoT
desbordante-core
@DesbordanteDesbordante is a high-performance data profiler that is capable of discovering many different patterns in data using various algorithms. It also allows to run data cleaning scenarios using these algorithms. Desbordante has a console version and an easy-to-use web application.
devlake
@apacheApache DevLake is an open-source dev data platform to ingest, analyze, and visualize the fragmented data from DevOps tools, extracting insights for engineering excellence, developer experience, and community growth.
cocoindex
@cocoindex-ioIncremental engine for long horizon agents 🌟 Star if you like it!
risingwave
@risingwavelabsEvent streaming platform for agentic AI. Continuously ingest, transform, and serve event streams in real time, at scale.
hudi
@apacheUpserts, Deletes And Incremental Processing on Big Data.
data-engineering-wiki
@data-engineering-communityThe best place to learn data engineering. Built and maintained by the data engineering community.
squid-sdk
@subsquidTypeScript ETL toolkit for indexing Ethereum, Solana, and Substrate data, sourced from SQD Network.
airbyte
@airbytehqOpen-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses, lakes, and AI applications. Both self-hosted and Cloud.
hamilton
@apacheApache Hamilton helps data scientists and engineers define testable, modular, self-documenting dataflows, that encode lineage/tracing and metadata. Runs and scales everywhere python does.
Daft
@Eventual-IncHigh-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
mlops-python-package
@fmindA comprehensive Python package template to kickstart and standardize your MLOps initiatives and data pipelines.
dbt
@dbt-labsdbt enables data analysts and engineers to transform their data using the same practices that software engineers use to build applications.
debezium
@debeziumChange data capture for a variety of databases. Please log issues at https://github.com/debezium/dbz/issues.
jitsu
@jitsucomJitsu is an open-source Segment alternative. Fully-scriptable data ingestion engine for modern data teams. Set-up a real-time data pipeline in minutes, not days
meltano
@meltanoMeltano: the declarative code-first data integration engine that powers your wildest data and ML-powered product ideas. Say goodbye to writing, maintaining, and scaling your own API integrations.
olake
@datazip-incOLake - Fastest Databases, Kafka & S3 Replication to Apache Iceberg with Table optimization (Called OLake Fusion). ⚡ Efficient, quick and scalable data ingestion for real-time analytics. Supported sources : Postgres, MongoDB, MySQL, Oracle, MSSql, DB2, Kafka, S3.
pgsync
@toluainaPostgres, MySQL, or MariaDB to Elasticsearch/OpenSearch sync
egeria
@odpiEgeria core
DataFlow-Engine
@risesoft-y9数据流引擎是一款面向数据集成、数据同步、数据交换、数据共享、任务配置、任务调度的底层数据驱动引擎。数据流引擎采用管执分离、多流层、插件库等体系应对大规模数据任务、数据高频上报、数据高频采集、异构数据兼容的实际数据问题。
maestro
@NetflixMaestro: Netflix’s Workflow Orchestrator
ape-dts
@apecloudApeCloud's Data Transfer Suite, written in Rust. Provides ultra-fast data replication between MySQL, PostgreSQL, Redis, MongoDB, Kafka and ClickHouse, ideal for disaster recovery (DR) and migration scenarios.
aws-sdk-pandas
@awspandas on AWS - Easy integration with Athena, Glue, Redshift, Timestream, Neptune, OpenSearch, QuickSight, Chime, CloudWatchLogs, DynamoDB, EMR, SecretManager, PostgreSQL, MySQL, SQLServer and S3 (Parquet, CSV, JSON and EXCEL).
hop
@apacheHop Orchestration Platform
ReplicaDB
@osalvadorReplicaDB is open source tool for database replication, designed for efficiently transferring bulk data between relational and non-relational databases
od
@kokesČeská otevřená data
apple-notes-exporter
@kzaremskiMacOS app written in Swift that bulk exports Apple Notes (including iCloud Notes) to a multitude of formats preserving note folder structure.
pudl
@catalyst-cooperativeThe Public Utility Data Liberation Project provides analysis-ready energy system data to climate advocates, researchers, policymakers, and journalists.
odd-platform
@opendatadiscoveryFirst open-source data discovery and observability platform. We make a life for data practitioners easy so you can focus on your business.
dagster
@dagster-ioAn orchestration platform for the development, production, and observation of data assets.
Addax
@wgzhaoActively maintained successor to Alibaba DataX — a fast, versatile, open-source ETL tool for 20+ RDBMS and NoSQL data sources.
datacontract-cli
@datacontractEnforce Data Contracts
gspread-pandas
@ParadigmllcRead and write Google Sheets as pandas DataFrames — column-matched appends, real dtypes, and layout detection for sheets that don't start at A1.
spark-excel
@nightscapeA Spark plugin for reading and writing Excel files
tis
@datavaneSupport agile Ontology DataOps Based on Flink, DataX and Flink-CDC with Web-UI
open-data-contract-standard
@bitol-ioHome of the Open Data Contract Standard (ODCS).
rocky
@rocky-dataA SQL transformation engine that type-checks your whole pipeline and catches breaking changes before they run — branches, replay, column-level lineage, compile-time contracts, per-model cost. Adapters: Databricks, Snowflake, BigQuery, DuckDB. Single static Rust binary. Apache 2.0.
riko
@nerevuA Python stream processing engine modeled after Yahoo! Pipes
monopoly
@benjamin-awdMonopoly is a Python library & CLI that converts bank statement PDFs to CSV.
pyjanitor
@pyjanitor-devsClean APIs for data cleaning. Python implementation of R package Janitor
conduit
@ConduitIOConduit streams data between data stores. Kafka Connect replacement. No JVM required.
filesql
@nao1215loads CSV, TSV, LTSV, JSON, JSONL, Parquet, XLSX, ACH, and Fedwire files into SQLite; includes prep and frame for cleanup and in-memory transforms
sling-cli
@slingdata-ioSling is a CLI tool that extracts data from a source storage/database and loads it in a target storage/database.
koheesio
@Nike-IncPython framework for building efficient data pipelines. It promotes modularity and collaboration, enabling the creation of complex pipelines from simple, reusable components.
dbt-databricks
@databricksA dbt adapter for Databricks.
spark
@dataflintDrop-in replacement for Apache Spark UI
chunjun
@DTStackA data integration framework
soda-core
@sodadataData Contracts engine for the modern data stack. https://www.soda.io
zerocode
@authorjappszerocode-tdd is a community-developed, free, open-source, outcome-driven automated testing framework for Data Pipelines, ETL, REST API, Kafka(Data Streams), Databases and Load scenarios. all defined in simple JSON or YAML — with zero coding.
datavines
@datavaneKnow your data better!Datavines is Next-gen Data Observability Platform, support metadata manage and data quality.
bacalhau
@bacalhau-projectCommunity-driven, simple, yet powerful framework for fast, cost-effective distributed Compute over Data.
extractor
@lightfeedUse LLMs to robustly extract web data
CommonCoreOntologies
@CommonCoreOntologyThe Common Core Ontology Repository holds the current released version of the Common Core Ontology suite.
snowpark-python
@snowflakedbSnowflake Snowpark Python API
dbt-trino
@starburstdataThe Trino (https://trino.io/) adapter plugin for dbt (https://getdbt.com)
dbt-sqlserver
@dbt-msftdbt adapter for SQL Server and Azure SQL
lakeFS
@treeverselakeFS - Data version control for your data lake | Git for data
fluvio
@fluvio-community🦀 event stream processing for developers to collect and transform data in motion to power responsive data intensive applications.
recce
@DataRecceThe data-validation toolkit for enhanced dbt (data build tool) PR review
DataMate
@ModelEngine-GroupDataMate is an enterprise-level data processing platform designed for model fine-tuning and RAG retrieval.
Flowfile
@EdwardvaneechoudFlowfile is a visual ETL tool and Python library combining drag-and-drop workflows with Polars dataframes. Build data pipelines visually, define flows programmatically with a Polars-like API, and export to standalone Python code. Perfect for fast, intuitive data processing from development to production.
kafka-connect-file-pulse
@streamthoughts🔗 A multipurpose Kafka Connect connector that makes it easy to parse, transform and stream any file, in any format, into Apache Kafka
DataEngineeringPilipinas
@ogbinarData Engineering Pilipinas is a community for data engineers, data analysts, data scientists, developers, AI / ML engineers, and users of closed and open source data tools and methods / techniques in the Philippines. Data Engineering Pilipinas is a PyData group.
dflib
@dflibIn-memory Java DataFrame library
pathway
@pathwaycomPython ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG.
dbt-coves
@datacovesCLI tool for dbt users to simplify creation of staging models (yml and sql) files
nomenklatura
@opensanctionsFramework and command-line tools for integrating FollowTheMoney data streams from multiple sources
CloudTAK
@dfpc-coeTAK Compatible, browser based Common Operation Picture & Situational Awareness tool
qsv
@dathereBlazing-fast Data-Wrangling toolkit
zdh_web
@zhaoyachao大数据采集,抽取平台,zdh_web是zdh系列服务的可视化管理平台,包含数据采集,调度,权限,审批流,私域营销等模块
growthbook
@growthbookOpen Source Feature Flags, Experimentation, and Product Analytics
wexflow
@aelassasWorkflow Automation Engine
automate-dv
@Datavault-UKA free to use dbt package for creating and loading Data Vault 2.0 compliant Data Warehouses (powered by dbt, an open source data engineering tool, registered trademark of dbt Labs)
lakehouse-engine
@adidasThe Lakehouse Engine is a configuration driven Spark framework, written in Python, serving as a scalable and distributed engine for several lakehouse algorithms, data flows and utilities for Data Products.
morph-kgc
@morph-kgcPowerful RDF knowledge graph generation with RML mappings
hflow
@Hebbian-RoboticsSDK for robotics teams to verify the quality of their data used for AI model training.
steampipe-plugin-aws
@turbotUse SQL to instantly query AWS resources across regions and accounts. Open source CLI. No DB required.
starflow
@starlake-aiDeclarative text based tool for data analysts and engineers to extract, load, transform and orchestrate their data pipelines.
skills
@dagster-ioA collection of AI skills for working with Dagster
harmonypy
@slowkow🎼 Integrate multiple high-dimensional datasets with fuzzy k-means and locally linear adjustments.
mloda
@mloda-aimloda.ai - Open Data Access for AI and ML. Plugin-based. Traceable. Framework-agnostic.
streamable
@ebonnalsync/async iterable streams for Python
connect
@redpanda-dataFancy stream processing made operationally mundane
awesome-engineering-articles
@ashishps1A curated collection of 300+ engineering blog articles from top tech companies. Learn how the best engineering teams solve real-world problems at scale.
sqrl
@DataSQRLAgentic Data Engineering Harness for building data pipelines, data products, data APIs, and data lakes autonomously
xonsh
@xonsh🐚 Python-powered shell. Full-featured, cross-platform and AI-friendly.
superglue
@superglue-aisuperglue (YC W25) builds integrations and tools from natural language. Get production-grade tools for long tail and enterprise systems.
zer0share
@zer0quantA 股、期货、期权数据本地化管道:Tushare Pro 拉取 → Parquet 分区存储 → DuckDB 本地查询,支持增量同步与定时调度
ingestr
@bruin-dataingestr is a CLI tool to copy data between any databases with a single command seamlessly.
usaspending-api
@fedspendingtransparencyServer application to serve U.S. federal spending data via a RESTful API
mcp-cn-commerce
@TonyWang-hub中国电商 MCP 连接器:8 个平台、155 个已注册工具。开源标准版,含能力矩阵与实际工程验收记录;真实商家联调状态逐项公开。Chinese commerce MCP connectors with documented capabilities and validation status.
datajoint-python
@datajointRelational Workflows: where database schemas define executable data pipelines.
snowplow
@snowplowThe leader in Customer Data Infrastructure
quilt
@quiltdataQuilt is a Scientific Data Management Platform on AWS that helps teams and AI find, trust, and reuse data through deeply versioned, context-rich data packages.
beginner_de_project
@josephmachadoBeginner data engineering project - batch edition
databricks_bootcamp_2026
@DataWithBaraaEnd-to-end Data Lakehouse project built on Databricks, following the Medallion Architecture (Bronze, Silver, Gold). Covers real-world data engineering and analytics workflows using Spark, PySpark, SQL, Delta Lake, and Unity Catalog. Designed for learning, portfolio building, and job interviews.
wingfoil
@wingfoil-ioultra low latency graph based stream processing framework
paperetl
@neuml📄 ⚙️ ETL processes for medical and scientific papers
cnpj-data-pipeline
@caiopizzolPipeline open-source que baixa e processa os dados da Receita Federal para PostgreSQL
pentaho-kettle
@pentahoPentaho Data Integration ( ETL ) a.k.a Kettle
sql-translator
@whoiskatrinSQL Translator is a tool for converting natural language queries into SQL code using artificial intelligence. This project is 100% free and open source.
amphi-etl
@amphi-aivisual data prep powered by python
ChoETL
@CinchooETL framework for .NET (Parser / Writer for CSV, Flat, Xml, JSON, Key-Value, Parquet, Yaml, Avro formatted files)
extract
@ICIJA cross-platform command line tool for parallelised content extraction and analysis.
ethereum-etl
@blockchain-etlPython scripts for ETL (extract, transform and load) jobs for Ethereum blocks, transactions, ERC20 / ERC721 tokens, transfers, receipts, logs, contracts, internal transactions. Data is available in Google BigQuery https://goo.gl/oY5BCQ
flow
@estuary🌊 Continuously synchronize the systems where your data lives, to the systems where you _want_ it to live, by managing your data flows with Estuary. 🌊
recap
@gabledataWork with your web service, database, and streaming schemas in a single format.
public-datasets-pipelines
@GoogleCloudPlatformCloud-native, data onboarding architecture for Google Cloud Datasets
data-science-on-gcp
@GoogleCloudPlatformSource code accompanying book: Data Science on the Google Cloud Platform, Valliappa Lakshmanan, O'Reilly 2017
Data-engineering-nanodegree
@Flor91Projects done in the Data Engineering Nanodegree by Udacity.com
go-etl
@Breeze0806go-etl is a toolset for data extraction, transformation and loading.
incubator-graphar
@apacheAn open source, standard data file format for graph data storage and retrieval.
opendbt
@memiisoMake dbt great again! Extend dbt with plugins, local docs and custom adapters — fast, safe, and developer-friendly
flowcraft
@gorangoA lightweight workflow engine
pg_background
@vibhorkumProduction-grade PostgreSQL extension to execute arbitrary SQL in background worker processes — with async execution, autonomous transactions, cookie-protected handles, cancellation, progress reporting, and observability.
Exchangis
@WeBankFinTechExchangis is a lightweight,highly extensible data exchange platform that supports data transmission between structured and unstructured heterogeneous data sources
databricks-code-practice
@jrlasakPractice Databricks coding skills with hands-on exercises. Import into Databricks Free Edition, write code, run assertions, check pass/fail. Covers Delta Lake, Spark SQL, PySpark, Auto Loader, medallion architecture, window functions, and more.
pypi-duck-flow
@mehd-ioend-to-end data engineering project to get insights from PyPi using python, duckdb, MotherDuck
DataCleaner
@datacleanerThe premier open source Data Quality solution
redun
@insitroYet another redundant workflow engine
bulk-writer
@jbogardProvides guidance for fast ETL jobs, an IDataReader implementation for SqlBulkCopy (or the MySql or Oracle equivalents) that wraps an IEnumerable, and libraries for mapping entites to table columns.
pglogical
@2ndQuadrantLogical Replication extension for PostgreSQL 17, 16, 15, 14, 13, 12, 11, 10, 9.6, 9.5, 9.4 (Postgres), providing much faster replication than Slony, Bucardo or Londiste, as well as cross-version upgrades.
qData
@qiantongtechqData is an open-source data governance and data development platform that integrates ETL, data development, metadata management, data quality, data assets, API services, and AI-powered data Q&A.
gusty
@pipeline-toolsMaking DAG construction easier
airflow-dbt-python
@tomasfariasA collection of Airflow operators, hooks, and utilities to elevate dbt to a first-class citizen of Airflow.
complete-dbt-bootcamp-zero-to-hero
@zoltanctothSupplementary Materials for the The Complete dbt (Data Build Tool) Bootcamp Udemy course
tributary
@1kbgzStreaming reactive and dataflow graphs in Python
radient
@fzliuRadient turns many data types (not just text) into vectors for similarity search, RAG, regression analysis, and more.
go-streams
@reugnA lightweight stream processing library for Go
seatunnel-web
@apacheSeaTunnel is a distributed, high-performance data integration platform for the synchronization and transformation of massive data (offline & real-time).
metorikku
@YotpoLtdA simplified, lightweight ETL Framework based on Apache Spark
meilisync
@long2iceRealtime sync data from MySQL/PostgreSQL/MongoDB to Meilisearch
nichenetr
@saeyslabNicheNet: predict active ligand-target links between interacting cells
practical-data-engineering
@ssp-dataPractical Data Engineering: A Hands-On Real-Estate Project Guide
dagster-open-platform
@dagster-ioDagster Labs' open-source data platform, built with Dagster.
modern-polars
@kevinheaveyCode and data for the Modern Polars book
fhir-data-pipes
@ohs-foundationA collection of tools for extracting FHIR resources and analytics services on top of that data.
metl
@jumpmindincMetl is a simple, web-based integration platform that allows for several different styles of data integration including messaging, file based Extract/Transform/Load (ETL), and remote procedure invocation via Web Services. Read more at www.jumpmind.com/products/metl/overview
jaffle-shop
@dbt-labs🥪🦘 An open source sandbox project exploring dbt workflows via a fictional sandwich shop's data.
every-single-day-i-tldr
@sderosiauxA daily digest of the articles or videos I've found interesting, that I want to share with you.
dcs-core
@datachecksOpen Source Data Quality Monitoring.
sql-ultimate-course
@DataWithBaraaThe most comprehensive SQL guide from a real-world expert! Learn everything from basics to advanced queries, optimizations, and real-world SQL
PyAirbyte
@airbytehqPyAirbyte brings the power of Airbyte to every Python developer. Powers the Airbyte Cloud Replication MCP.
kuwala
@kuwala-ioKuwala is the no-code data platform for BI analysts and engineers enabling you to build powerful analytics workflows. We are set out to bring state-of-the-art data engineering tools you love, such as Airbyte, dbt, or Great Expectations together in one intuitive interface built with React Flow. In addition we provide third-party data into data science models and products with a focus on geospatial data. Currently, the following data connectors are available worldwide: a) High-resolution demographics data b) Point of Interests from Open Street Map c) Google Popular Times
sql-data-warehouse-project
@DataWithBaraaA comprehensive guide to building a modern data warehouse with SQL Server, including ETL processes, data modeling, and analytics.
dozer
@getdozerDozer is a real-time data movement tool that leverages CDC from various sources and moves data into various sinks.
CQL
@CategoricalDataCategorical Query Language IDE
orbital
@orbitalapiOrbital automates integration between data sources (APIs, Databases, Queues and Functions). BFF's, API Composition and ETL pipelines that adapt as your specs change.
aitrados-api
@aitradosOHLC,news,economic event restfull api and WebSocket API for specifically designed for AI quantitative trading/training.Multiple Timeframes,Multiple-Symbols-Multiple-Timeframes
monstache
@rwynna go daemon that syncs MongoDB to Elasticsearch in realtime. you know, for search.
efficient_data_processing_spark
@josephmachadoCode for "Efficient Data Processing in Spark" Course
hellodata-be
@kanton-bernThe Open-Source Enterprise Data Platform in a single Portal
ethereum-etl-airflow
@blockchain-etlAirflow DAGs for exporting, loading, and parsing the Ethereum blockchain data. How to get any Ethereum smart contract into BigQuery https://towardsdatascience.com/how-to-get-any-ethereum-smart-contract-into-bigquery-in-8-mins-bab5db1fdeee
kiba
@thbarData processing & ETL framework for Ruby
diffsync
@networktocodeA utility library for comparing and synchronizing different datasets.
awesome-kafka
@infoslackA list about Apache Kafka
hudi-resources
@leesf汇总Apache Hudi相关资料
butterfree
@quintoandarA tool for building feature stores.
abc
@appbaseioPower of appbase.io via CLI, with nifty imports from your favorite data sources
goodreads_etl_pipeline
@san089An end-to-end GoodReads Data Pipeline for Building Data Lake, Data Warehouse and Analytics Platform.
memphis
@superstreamlabsMemphis.dev is a highly scalable and effortless data streaming platform
DataEngineeringProject
@damklisExample end to end data engineering project.
flupy
@oliriceFluent data pipelines for python and your shell
airbyte_serverless
@unyticsAirbyte made simple (no UI, no database, no cluster)
dataplane
@dataplane-appDataplane is an Airflow inspired unified data platform with additional data mesh and RPA capability to automate, schedule and design data pipelines and workflows. Dataplane is written in Golang with a React front end.
pyper
@pyper-devConcurrent Python made simple
omniparser
@jf-techomniparser: a native Golang ETL streaming parser and transform library for CSV, JSON, XML, EDI, text, etc.
etl2pcapng
@microsoftUtility that converts an .etl file containing a Windows network packet capture into .pcapng format.
yobulkdev
@yobulkdev🔥 🔥 🔥Open Source & AI driven Data Onboarding Platform:Free flatfile.com alternative
mara-pipelines
@maraA lightweight opinionated ETL framework, halfway between plain scripts and Apache Airflow
harmony
@immunogenomicsFast, sensitive and accurate integration of single-cell data with Harmony
setl
@SETL-FrameworkA simple Spark-powered ETL framework that just works 🍺
getting-started
@singer-ioThis repository is a getting started guide to Singer.
klio
@spotifySmarter data pipelines for audio.
yuniql
@rdagumampanFree and open source schema versioning and database migration made natively with .NET/6. NEW THIS MAY 2022! v1.3.15 released!
bitcoin-etl
@blockchain-etlETL scripts for Bitcoin, Litecoin, Dash, Zcash, Doge, Bitcoin Cash. Available in Google BigQuery https://goo.gl/oY5BCQ
flock
@flock-labFlock: A Low-Cost Streaming Query Engine on FaaS Platforms
spark-alchemy
@swoop-incCollection of open-source Spark tools & frameworks that have made the data engineering and data science teams at Swoop highly productive
dud
@kevin-hanselmanA lightweight CLI tool for versioning data alongside source code and building data pipelines.
smooks
@smooksAn extensible Java framework for building event-driven applications that break up XML and non-XML data into chunks for data integration
data-story
@ajthinkingA visual process builder
prefect-dataplatform
@anna-gellerExample repository showing how to build a data platform with Prefect, dbt and Snowflake
active_workflow
@automaticmodePolyglot workflows without leaving the comfort of your technology stack.
NeumAI
@NeumTryNeum AI is a best-in-class framework to manage the creation and synchronization of vector embeddings at large scale.
SmartCode
@dotnetcoreSmartCode = IDataSource -> IBuildTask -> IOutput => Build Everything!!!