dltHub is looking for a freelance help in the following repos:
- https://github.com/dlt-hub/dlt - https://github.com/dlt-hub/verified-sources
Please look at the issues, README and CONTRIBUTE guides - those are the tasks you’ll be working on. We will ensure you are onboarded and will give you comprehensive reviews of your code. We expect that you can work with us a minimum 20 hours/week and ideally you will be flexible to do some more upon request.
dlt is an open-source library that automatically creates datasets out of messy, unstructured data sources. You can use the library to move data from about anywhere into most of well-known SQL and vector stores, data lakes, storage buckets, or local engines like DuckDB. It automates many cumbersome data engineering tasks and can be handled by anyone who knows Python. You can read more about us here: https://news.ycombinator.com/item?id=37999527
---
Your Task and Responsibilities: - Contribute to dlt, including code, tests, and documentation - Maintain the open source project with the team (e.g., review PRs, resolve issues, talk with community contributors, etc.)
Who You Are: - You really like Python and are fluent in writing Python code (e.g., Python typing, unit testing, writing docstrings, etc.) - You are interested in the Python ecosystem (i.e., popular libraries, tools, Python internals, PEPs, how Python is used outside of software engineering, etc.) - You have some experience with databases and data warehouses (i.e., you understand the relational data model, transactions, concurrency, etc.) You are familiar with GitHub workflows (e.g., pull requests, code reviews, CI/CD services, etc.)
Nice to Have: - You might have experience with data engineering (e.g., building data pipelines, dataset modeling, enabling others to use the data, etc.) -You might have experience with machine learning (e.g., the toolset, the workflows, practical applications, etc.)
If the role sounds interesting, please apply here: https://apply.workable.com/dlthub/j/4C5D016FF5/apply/