Growth through diversity, equity, and inclusion. As an ethical business, we do what is right — including ensuring equal opportunities and fostering a safe, respectful workplace for each of us. We believe diversity fuels both personal and business growth. We're committed to building an inclusive community where all our people thrive regardless of their backgrounds, identities, or other personal characteristics.
What You'll Be Doing:
- Building data pipelines to ingest data from various sources such as databases, APIs, or streaming platforms. Integrating and transforming data to ensure its compatibility with the target data model or format
- Designing and optimizing data storage architectures, including data lakes, data warehouses, or distributed file systems. Implementing techniques like partitioning, compression, or indexing to optimize data storage and retrieval. Identifying and resolving bottlenecks, tuning queries, and implementing caching strategies to enhance data retrieval speed and overall system efficiency.
- Designing and implementing data models that support efficient data storage, retrieval, and analysis. Collaborating with data scientists and analysts to understand their requirements and provide them with well-structured and optimized data for analysis and modeling purposes.
- Collaborating with cross-functional teams including data scientists, analysts, and business stakeholders to understand their requirements and provide technical solutions. Communicating complex technical concepts to non-technical stakeholders in a clear and concise manner.
- Independence and responsibility for delivering a solution
- Ability to work under Agile and Scrum development methodologies
- Train and mentor junior data engineers, providing guidance and knowledge transfer
- Designing and implementing data processing systems on Azure platform. This involves writing efficient and scalable code to process, transform, and clean large volumes of structured and unstructured data.
- Building data pipelines to ingest data from various sources such as databases, APIs, or streaming platforms.
- Integrating and transforming data to ensure its compatibility with the target data model or format.
- Identifying and resolving bottlenecks, tuning queries, and implementing caching strategies to enhance data retrieval speed and overall system efficiency.
- Collaborating with cross-functional teams including data scientists, analysts, and business stakeholders to understand their requirements and provide technical solutions.
What We're Looking For:
-
At least 4 years of experience as a Data Engineer working with GCP cloud-based infrastructure & systems.
-
Deep knowledge of Google Cloud Platform and cloud computing services.
-
Extensive experience in design, build, and deploy data pipelines in the cloud, to ingest data from various sources like databases, APIs or streaming platforms.
-
Proficient in database management systems such as SQL (Big Query is a must), NoSQL. Candidate should be able to design, configure, and manage databases to ensure optimal performance and reliability.
-
Programming skills (SQL, Python, other scripting).
-
Proficient in data modeling techniques and database optimization. Knowledge of query optimization, indexing, and performance tuning is necessary for efficient data retrieval and processing.
-
Knowledge of at least one orchestration and scheduling tool (Airflow is a must).
-
Experience with data integration tools and techniques, such as ETL and ELT Candidate should be able to integrate data from multiple sources and transform it into a format that is suitable for analysis.
-
Knowledge of modern data transformation tools (such as DBT, Dataform).
-
Excellent communication skills to effectively collaborate with cross-functional teams, including data scientists, analysts, and business stakeholders. Ability to convey technical concepts to non-technical stakeholders in a clear and concise manner.
-
Ability to actively participate/lead discussions with clients to identify and assess concrete and ambitious avenues for improvement.
-
Tools knowledge: Git, Jira, Confluence, etc.
-
Open to learn new technologies and solutions.
-
Experience in multinational environment and distributed teams.
What Will Set You Apart:
- Certifications in big data technologies or/and cloud platforms.
- Experience with BI solutions (e.g. Looker, Power BI, Tableau).
- Experience with ETL tools: e.g. Talend, Alteryx
- Experience with Apache Spark, especially in GCP environment.
- Experience with Databricks.
- Experience with Azure cloud-based infrastructure & systems.
Missing one or two of these qualifications? We still want to hear from you! If you bring a positive mindset, we'll provide an environment where you feel valued and empowered to learn and grow.
We offer:
- Stable employment. On the market since 2008, 1800+ talents currently on board in 7 global sites.
- Full-time position with work contract.
- Medical Insurance.
- Grocery Coupons.
- Saving fund.
- 30 days of Christmas bonus.
- Remote work bonus.
- Profit sharing.
- 50% vacation premium.
- 100% remote.
- Flexibility regarding working hours.
- Comprehensive online onboarding program with a “Buddy” from day 1.
- Cooperation with top-tier engineers and experts.
- Unlimited access to the Udemy learning platform from day 1.
- Certificate training programs. Lingarians earn 500+ technology certificates yearly.
- Upskilling support. Capability development programs, Competency Centers, knowledge sharing sessions, community webinars, 110+ training opportunities yearly.
- Grow as we grow as a company. 76% of our managers are internal promotions.
- A diverse, inclusive, and values-driven community.
- Autonomy to choose the way you work. We trust your ideas.
- Create our community together. Refer your friends to receive bonuses.
- Activities to support your well-being and health.
- Plenty of opportunities to donate to charities and support the environment.