I will write python data validation and quality checks for your etl pipeline

Taiwan

I speak Chinese, English

AI Voice Cloning in Mandarin, Automation and Data Pipelines

Software engineer & data pipeline builder from Taiwan. I run a one-person AI studio (SonarCue) specializing in AI voice cloning with natural Mandarin/Chinese output, plus automation and data pipelines...
About this Gig

Stop silent data errors: automated Python + pytest checks that catch the failures your logs call "success".


Pipelines rarely crash they quietly return wrong data, like an API returning an empty list while the job logs success. I build the checks that catch those patterns.


WHAT YOU GET

- A findings report on issues in scope, with evidence and repro steps

- Custom checks in Python (pandas + pytest), full source code, yours to keep

- Premium: a GitHub Actions workflow for one repo, plus a failure runbook


CHECKS COVER: freshness, completeness, schema drift, cross-source consistency, failure-state accuracy.


SCOPE

- 1 Data Source = 1 CSV/JSON file, 1 table or view, or 1 API endpoint with one schema. Check counts are totals per order.

- Revisions: one round on the originally agreed data. New sources or rules: a Custom Offer.


Built from 32 documented real failures from running my own data systems sample cases on GitHub; a Sample Findings Report is in the gallery.


NO CALLS everything in writing on Fiverr. No DB credentials for Basic/Standard: send a data sample or a public endpoint URL.


A software engineering service I don't advise on what your data means for your business.

Langugae:

Chinese (Simplified)

English

Technical expertise:

Other

Expertise:

Data Pipelines

Data Governance

Industry:

Data analytics

Financial services