Wearables Data Engine
13,000+ lines of Python normalizing Rook and Spike payloads from Apple Health, Health Connect, Samsung Health, and Whoop into daily metrics and activity points. 10,000+ records a day, 1M+ accumulated.
Data & Automation Engineer
I build the pipelines, automation, and dashboards behind a health-tech platform processing 10,000+ wearable records a day. Based in Cairo, open to Gulf relocation.
GitHub·LinkedIn·arlabib1192002@gmail.com·+20 155 546 4744·Resume ↓

Three years into building production systems at BytePlus Life, a Cairo health-tech startup: a wearables data engine that has accumulated over a million records, a 63-endpoint operations API, a persona-based notification engine in Arabic and English, and the live operations dashboard the company runs on. Alongside it: a B.Sc. in Artificial Intelligence from Cairo University, ECPC 2023, and a stream of shipped side projects.
BytePlus Life (Byte+) · Jan 2024 to Present · code private, systems live
13,000+ lines of Python normalizing Rook and Spike payloads from Apple Health, Health Connect, Samsung Health, and Whoop into daily metrics and activity points. 10,000+ records a day, 1M+ accumulated.
63 REST endpoints and 17 tables covering the full B2B client lifecycle, tested by 200+ checks against in-process Postgres plus 60+ live staging end-to-end checks.
React 19 control center for onboarding, contracts, and seasons. Re-hosted on S3 + CloudFront after diagnosing local ISPs black-holing Netlify’s edge IPs; 110 KB lighter first paint. Live at ops.bytepluslife.com (internal login).
Persona-based engagement engine: 23 contextual detectors, bilingual Arabic and English template banks, corporate frequency guardrails, Ramadan-aware scheduling.
FastAPI microservice computing weighted health scores and biological-age deltas from blood, InBody, and physiotherapy assessments.
11-step pipeline with a Streamlit control plane; onboarded 10+ corporate clients, then ported into the backend as production endpoints.
A health-tech data domain rebuilt on an open stack: dbt (15 models, SCD2 snapshots), Airflow DAGs, Kafka streaming, PySpark batch scoring, and an A/B analysis with power calculations. 7.2M rows of seeded synthetic data, 8 architecture decision records, CI green.


Chrome side-panel YouTube research copilot: a multi-strategy transcript pipeline built to survive YouTube’s proof-of-origin defenses, provider-agnostic AI routing (Gemini, OpenAI, Anthropic, Ollama), local-first IndexedDB workspaces.
WhatsApp Business Cloud webhook and agent inbox: raw-byte signature verification, idempotent ingestion, signed media URLs, and a machine-to-machine send API. Runs a production WhatsApp line.
Open to Data Engineer, Data Analyst, and Backend roles in Cairo or the Gulf (can relocate within 30 days).