Data Engineer
Role Description
As a Data Engineer, you will design, build, and maintain the data infrastructure and pipelines that power analytics and decision-making across the organization. You will work on the full data lifecycle – from ingestion and transformation to storage, orchestration, and delivery of reliable, high-quality datasets.
You will collaborate closely with data scientists, analysts, software engineers, and business stakeholders to understand data needs and to build scalable, well-governed data platforms that teams can trust and depend on.
Key Responsabilities
- Design, build, and maintain scalable data pipelines (ETL/ELT) to ingest data from multiple sources.
- Develop and optimize data models, warehouses, and lakes to support analytics and reporting needs.
- Ensure data quality, consistency, and integrity across systems through validation, testing, and monitoring.
- Automate data workflows and orchestration using modern pipeline and scheduling tools.
- Collaborate with data scientists and analysts to make data readily accessible and well-structured for their use cases.
- Optimize the performance, scalability, and cost-efficiency of data infrastructure.
- Implement data governance, security, and privacy best practices across pipelines and storage systems.
- Troubleshoot and resolve data pipeline issues, ensuring high availability and reliability.
- Document data architecture, pipelines, and processes for maintainability and knowledge sharing.
- Stay current with emerging data engineering tools and practices, proposing their adoption when relevant.
Job Qualifications
- Bachelor’s or Master’s degree in Computer Science, Engineering, Information Systems, or a related field.
- 3+ years of experience in a Data Engineer or similar role.
- Strong understanding of data modeling, warehousing, and distributed systems concepts.
- Proven experience building and maintaining production-grade data pipelines.
- Strong problem-solving skills and attention to data quality and detail.
- Good communication skills, with the ability to work closely with technical and business stakeholders.
- Fluency in English (written and spoken); additional language is a plus.
- Ability to work independently and as part of a distributed/remote team.
Main Tech Skills
- Programming Languages: Python and/or Scala/Java; strong SQL skills.
- Data Pipelines & Orchestration: Apache Airflow, dbt, Luigi, or similar.
- Big Data Processing: Apache Spark, Hadoop, Kafka.
- Data Warehousing: Snowflake, BigQuery, Redshift, or Synapse.
- Cloud Platforms: AWS (Glue, S3, EMR), Azure (Data Factory, Synapse), or GCP (Dataflow, BigQuery).
- Databases: SQL and NoSQL databases (PostgreSQL, MySQL, MongoDB, Cassandra).
- Containerization & Orchestration: Docker, Kubernetes.
- CI/CD & Version Control: Git, GitHub Actions/GitLab CI/Jenkins.
- Infrastructure as Code: Terraform or CloudFormation (a plus).
- Monitoring & Data Quality: Great Expectations, Monte Carlo, or similar tools (a plus).
Be Bold · Work Smart · Change Tomorrow
We started this company with a simple observation: most organizations don’t struggle because they lack technology — they struggle because solutions are too complex, disconnected from reality, or hard to sustain over time. Data and AI can create enormous value. But only when they are designed with clear intent, solid foundations, and a realistic understanding of how organizations actually operate. We’re a young, fast-growing company with plenty of opportunities to learn, evolve, and build something meaningful together.