Job Title – Data Architect
Key Responsibilities:
Design, develop, and maintain scalable data pipelines and architecture for data integration and transformation.
Collaborate with data scientists, analysts, and other stakeholders to understand data requirements and ensure architecture aligns with business goals.
Utilize Python and PySpark to process, transform, and analyze large volumes of structured and unstructured data.
Define and enforce data modeling standards and best practices.
Ensure the security, reliability, and performance of data systems.
Work with cloud-based data platforms (e.g., AWS, Azure, GCP) and big data technologies as required.
Develop and maintain metadata, data catalogs, and data lineage documentation.
Monitor and troubleshoot performance issues related to data pipelines and architecture.
Required Skills and Qualifications:
Bachelor's or master’s degree in computer science, Information Technology, or a related field.
5 to 8 years of hands-on experience in Data Architect roles.
Strong proficiency in Python and/or PySpark for data transformation and ETL processes.
Experience with distributed data processing frameworks like Apache Spark.
Experience working with relational and NoSQL databases (e.g., PostgreSQL, Cassandra, MongoDB).
Familiarity with data governance, security, and compliance principles.
Experience with CI/CD pipelines, version control (e.g., Git), and Agile methodologies.