Preparing for system design interviews?  Try bugzed.com →

datasieve: Flexible Data Pipeline

JSON →
library 0.1.9 ·python
verified Jun 28, 2026

The `datasieve` package provides a flexible data pipeline inspired by scikit-learn's Pipeline, but with enhanced capabilities to manipulate `y` (target) and `sample_weight` arrays alongside `X` (features). This is particularly useful for tasks such as removing outliers across all associated data, removing feature columns based on arbitrary criteria, and handling dynamic feature renaming within the pipeline. The current version is 0.1.9, with releases occurring on an irregular, as-needed basis.

total hits 16
actors 5 distinct systems
last hit 16d ago AhrefsBot
GPTBot
4
Script
1
ByteDance
1
ChatGPT-User
1
Humans
5

top countries 🇺🇸 United States · VN · 🇨🇦 Canada · 🇸🇬 Singapore · 🇫🇷 France