[Job - 30760] Master Data Developer, Colombia
CiandtColombiaHomeoffice2d ago
TypeScriptPythonAWSTerraformDynamoDBMachine LearningAIGitLinuxApacheData EngineeringETL
Job description
[Job - 30760] Master Data Developer, Colombia at Ciandt.
About the role
Ciandt is looking for a skilled Data Developer to advance our data and analytics capabilities. You will design and maintain scalable cloud-native data pipelines while optimizing storage and processing for our enterprise clients.
Key facts
What you'll do
- Create and maintain ETL and ELT workflows to manage data movement within modern lake architectures.
- Use Python and PySpark to process large-scale datasets.
- Manage data storage partitioning using frameworks like Delta Lake or Apache Iceberg to control costs and improve query speed.
- Write and refine complex SQL, including window functions and aggregations.
- Move data workloads from traditional RDBMS systems to cloud-native environments.
- Use AWS services like Glue, Athena, Redshift, S3, Lambda, and EventBridge.
- Collaborate with DevOps teams to manage resources via Infrastructure as Code.
- Monitor pipeline performance and data quality through observability tools.
Requirements
- Proven experience building ETL processes and data pipelines within AWS.
- Advanced proficiency in Python and PySpark for distributed data tasks.
- Expert-level SQL skills, including experience migrating legacy RDBMS workloads.
- Hands-on experience with AWS Glue, Athena, and Redshift.
- Knowledge of Data Lake architecture and storage optimization strategies.
- Understanding of object-oriented programming to create reusable code.
- Proficiency with Git, Shell scripting, and Linux.
- Familiarity with monitoring and metric tracking.
- Advanced or fluent English language skills.
Nice to have
- Familiarity with open table formats like Delta Lake or Apache Iceberg.
- Experience with TypeScript.
- Practical use of Infrastructure as Code tools such as CloudFormation, CDK, or Terraform.
- Exposure to AWS services like SageMaker AI, ECS, RDS, DynamoDB, IAM, or EventBridge.
- Experience using Pandas for data analysis.
- Background in machine learning or AI-related data projects.
Skills & tools
- Python, PySpark, SQL, Git, Shell, Linux.
- AWS (Glue, Athena, Redshift, S3, Lambda, EventBridge).
- Data Lake architectures, ETL/ELT.
Practical notes
- Benefits include premium healthcare, meal vouchers, maternity and parental leave, mobile service subsidies, sick pay, life insurance, access to Ciandt University, Colombian holidays, and paid vacation.