We are looking for a Data Engineer to support scalable data solutions for a long-term contract opportunity in The Woodlands, Texas. This role focuses on designing and optimizing data pipelines, integrating large-scale data sources, and enabling reliable access to critical business information. The ideal candidate will bring strong hands-on experience with modern big data technologies and a practical approach to building efficient ETL workflows.<br><br>Responsibilities:<br>• Build, maintain, and enhance robust data pipelines to process large volumes of structured and unstructured information.<br>• Develop ETL workflows that transform raw data into reliable datasets for analytics, reporting, and operational use.<br>• Use Python and Apache Spark to engineer high-performance data processing solutions across distributed environments.<br>• Work with Apache Hadoop ecosystems to manage storage and support scalable data operations.<br>• Integrate streaming and event-driven data using Apache Kafka to improve data availability and timeliness.<br>• Monitor data workflows, troubleshoot processing issues, and implement improvements that increase reliability and efficiency.<br>• Collaborate with technical and business stakeholders to understand data needs and translate them into practical engineering solutions.<br>• Document pipeline architecture, data flow logic, and operational procedures to support maintainability and team knowledge sharing.