Responsibilities
- 5+ years working on the role professionally.
- Coding: Coding Python and use in data processing solutions and related data technologies like Pandas, and PySpark.
- Consume data from different sources like REST APIs.
- Work with relational and non-relational data stores (like: HBASE, Cassandra or MongoDB; S3, blobs).
- Design.
- Data Streams (Kafka, Kinesis, Flume,) and message queuing (SQS, SNS, RabbitMQ, etc).
- Ensure that the data model scales and enables high performance.
- Distributed data stores.
- Data/Stream processing (Spark, Flink, Hadoop).
- Data pipelines, data ingestion pipelines, scalable streaming data pipelines processing.
- ETL using solutions: Talend; Informatica; SQL Server Integration Services (SSIS).
- Data warehouse (Snowflake, Redshift, Hive).
- Implementation of data warehouse solutions, providing near real-time data to a variety of client systems.
- Using SQL databases to construct data storage.
- Reporting / BI, design, implementation, and enhancement of BI tool is a plus.
- Experience designing and implementing data applications and services on the public cloud, AWS, GCP, or Azure using PaaS platforms.
- Familiarity with data privacy regulations and best practices.
Requirements
- 5+ years working on the role professionally.
- Coding: Coding Python and use in data processing solutions and related data technologies like Pandas, and PySpark.
- Consume data from different sources like REST APIs.
- Work with relational and non-relational data stores (like: HBASE, Cassandra or MongoDB; S3, blobs).
- Design.
- Data Streams (Kafka, Kinesis, Flume,) and message queuing (SQS, SNS, RabbitMQ, etc).
- Ensure that the data model scales and enables high performance.
- Distributed data stores.
- Data/Stream processing (Spark, Flink, Hadoop).
- Data pipelines, data ingestion pipelines, scalable streaming data pipelines processing.
- ETL using solutions: Talend; Informatica; SQL Server Integration Services (SSIS).
- Data warehouse (Snowflake, Redshift, Hive).
- Implementation of data warehouse solutions, providing near real-time data to a variety of client systems.
- Using SQL databases to construct data storage.
- Experience designing and implementing data applications and services on the public cloud, AWS, GCP, or Azure using PaaS platforms.
- Familiarity with data privacy regulations and best practices.
Nice to Have
- Reporting / BI, design, implementation, and enhancement of BI tool is a plus.
- Looker; Power BI; Tableau.
- Experience with Data Visualization tools such as Tableau.
- Data processing architecture like Lambda Architecture.
- Indexing (SolR or ElasticSearch).
- Prior experience converting Jupyter Notebooks into production processes.
- Data pipeline and workflow management tools. Airflow.
- Data lakes.
Benefits
- 100% Remote Work: Enjoy the freedom to work from the location that helps you thrive. All it takes is a laptop and a reliable internet connection.
- Highly Competitive USD Pay: Earn an excellent, market-leading compensation in USD, that goes beyond typical market offerings.
- Paid Time Off: We value your well-being. Our paid time off policies ensure you have the chance to unwind and recharge when needed.
- Work with Autonomy: Enjoy the freedom to manage your time as long as the work gets done. Focus on results, not the clock.
- Work with Top American Companies: Grow your expertise working on innovative, high-impact projects with Industry-Leading U.S. Companies.
Work Arrangement
Remote (Worldwide)
Additional Information
- A Culture That Values You: We prioritize well-being and work-life balance, offering engagement activities and fostering dynamic teams to ensure you thrive both personally and professionally.
- Diverse, Global Network: Connect with over 600 professionals in 25+ countries, expand your network, and collaborate with a multicultural team from Latin America.
- Team Up with Skilled Professionals: Join forces with senior talent. All of our team members are seasoned experts, ensuring you're working with the best in your field.