{"owner":"alireza-heidarii","github":"https://github.com/alireza-heidarii","claimed":false,"inventory":[],"indexed":[{"repo":"alireza-heidarii/Real-Time-Data-Cleaning-Pipeline-for-Medical-and-Healthcare-Data","github":"https://github.com/alireza-heidarii/Real-Time-Data-Cleaning-Pipeline-for-Medical-and-Healthcare-Data","description":"A real-time data cleaning pipeline for medical and healthcare data using Apache Spark, SparkNLP, Spark Streaming, and Kafka.","language":"Python","stars":13,"topics":["data-cleaning","data-pipelines","data-preprocessing","healthcare-datasets","kafka","medical-data-analysis","natural-language-processing","parquet","pyspark","python"],"license":"Apache-2.0","category":"data-pipelines"}],"how_to_buy":"GET /r/alireza-heidarii/<repo> (Accept: application/json) for any listed repo here: tree, README, price and the checkout to pay (x402; rehearse first at its test twin, simulated money). Repos under 'indexed' are free: clone them from GitHub."}