Jobs and Careers
PL
Senior Data Engineer
PlootoUnited States, United StatesRemotefull_timeVerifiedPosted 26 Oct 2023
About the role
We are currently hiring full-time in the following US locations: California, Georgia, Massachusetts, Florida, and Washington. If you are located outside of these states, you can still be considered as an applicant, but the position will be on a contract basis.
About Plooto82% of small businesses fail due to poor management of cash flow. Our vision is to enable the advancement of Small and Medium-sized Businesses (SMBs) by developing the tools & insights they need to maximize their cash flow. Over 10,000 businesses and their finance teams trust Plooto to automate their financial processes so they can focus on reaching their true potential.
About the RoleAs a Senior Data Engineer, you will help us in the upgrade of our Data Platform. Currently, the data engineering team has built a Plooto Data Platform on Azure, leveraging Azure Data Factory to create data integrating pipelines as well as Spark to create ETL jobs. The Data Platform and the built data products/models support our obsession with data-driven decision-making.The ideal candidate is an experienced ETL pipeline builder and data wrangler who enjoys building and optimizing data systems. You will work with the existing data engineers, business analysts, software engineers, and other internal stakeholders on data initiatives and will ensure optimal data delivery throughout ongoing projects. The right candidate will be excited by the prospect of optimizing or even re-designing our company’s data architecture to support our next generation of products and data initiatives. Candidates must be self-directed and comfortable supporting the data needs of multiple teams, systems, and products.
About Plooto82% of small businesses fail due to poor management of cash flow. Our vision is to enable the advancement of Small and Medium-sized Businesses (SMBs) by developing the tools & insights they need to maximize their cash flow. Over 10,000 businesses and their finance teams trust Plooto to automate their financial processes so they can focus on reaching their true potential.
About the RoleAs a Senior Data Engineer, you will help us in the upgrade of our Data Platform. Currently, the data engineering team has built a Plooto Data Platform on Azure, leveraging Azure Data Factory to create data integrating pipelines as well as Spark to create ETL jobs. The Data Platform and the built data products/models support our obsession with data-driven decision-making.The ideal candidate is an experienced ETL pipeline builder and data wrangler who enjoys building and optimizing data systems. You will work with the existing data engineers, business analysts, software engineers, and other internal stakeholders on data initiatives and will ensure optimal data delivery throughout ongoing projects. The right candidate will be excited by the prospect of optimizing or even re-designing our company’s data architecture to support our next generation of products and data initiatives. Candidates must be self-directed and comfortable supporting the data needs of multiple teams, systems, and products.
What You’ll Do
- Architect and implement data products and ETL pipelines for internal reporting (embedded reporting, dashboarding, insights) in the medium term, and for machine learning (insights to recommendations) in the long term.
- Establish a comprehensive Data Warehouse by creating robust ingestion pipelines using Azure Data Factory and developing resilient streaming pipelines with Kafka and Flinks for real-time event processing.
- Partner with the Product Manager, analytics, and business teams to review and gather the data/reporting/analytics requirements and build trusted and scalable data models, data extraction processes, and data applications to help answer complex questions.
- Build high volume, distributed, and scalable data platform capabilities and microservices to create efficient and scalable data solutions to enable data access by applications via API.
- Utilize and advance continuous integration and deployment frameworks.
- Work with architecture & engineering leads to ensure quality solutions are implemented and engineering standard methodologies adhered to.
Your Background
- 10+ years of data engineering in a B2B/B2C data-driven environment.
- Experience in designing, building, and optimizing complex data pipelines, ETL processes, and data integration solutions.
- Strong knowledge of big data technologies such as Databricks, Hadoop, Spark, and Kafka, as well as expertise in distributed data processing frameworks.
- Hands-on experience with cloud platforms like AWS, Azure, or Google Cloud, including knowledge of cloud-based data services (e.g., AWS Glue, Azure Data Factory, Google Cloud Dataflow). Knowledge of real-time data processing using technologies like Apache Kafka, Apache Flink, or similar streaming frameworks.
- Proficiency in SQL and experience with various database systems (SQL Server, PostgreSQL, NoSQL databases) for data modeling, optimization, and management.
- Proficiency in programming languages commonly used in data engineering, such as Python, Java, Scala, or similar languages.
- Strong understanding of data warehousing concepts, data modeling techniques (star, snowflake schemas), and experience with data warehousing solutions (e.g., Redshift, Snowflake).
- Understanding of data governance practices, data lineage, data privacy, and the ability to implement security measures for data handling.
Bonus Points
- Experienced with Airflow or other job scheduling systems.
- Experience with building Restful APIs and containerizing them.
- Contributions to open-source data projects.
- Distributed systems design/implementation experiences.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s