feathr

Data pipeline management

A unified data and AI engineering platform for enterprise

Feathr – A scalable, unified data and AI engineering platform for enterprise

GitHub

2k stars
84 watching
260 forks
Language: Scala
last commit: over 2 years ago
Linked from 2 awesome lists

apache-sparkartificial-intelligenceazuredata-engineeringdata-qualitydata-sciencefeature-engineeringfeature-governancefeature-managementfeature-marketplacefeature-metadatafeature-platformfeature-storemachine-learningmlops

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
apache/tezA system that enables flexible data processing pipelines using a low-level engine for higher-level frameworks482
floomai/floomA platform for orchestrating and executing Generative AI pipelines, empowering developers to automate complex tasks37
combust/mleapEnables deployment of machine learning pipelines from Spark and Scikit-Learn to production1,506
apache/streampipesA toolbox for industrial data analytics and stream processing614
tenzir/tenzirA data pipeline engine designed to manage and process large volumes of security telemetry data at scale651
linkedin/brooklinA distributed system for streaming data between heterogeneous systems with high reliability and throughput at scale931
apache/sparkAn analytics engine designed to handle large-scale data processing and analysis40,170
stratio/spartaAn Apache Spark-based platform for building real-time analytics workflows with a focus on simplicity and extensibility.525
huo-ju/dfserverA distributed backend AI pipeline server for building and managing GPU clusters to run various AI models.349
kbrw/adapA data augmentation pipeline built on top of Elixir, designed to process data streams and apply transformation rules in real-time.16
nessos/streamsA lightweight library for building efficient data pipelines using functional programming concepts383
cloud-cv/evalaiA platform for comparing and evaluating AI and machine learning algorithms at scale1,779
kevin-hanselman/dudA lightweight tool for managing and versioning large data alongside source code in data pipelines184
databricks/tensorframesEnables manipulation of Apache Spark DataFrames using TensorFlow programs749
webankfintech/dataspherestudioA comprehensive platform for managing and developing data applications, providing tools for data exchange, analysis, visualization, and workflow management.3,100