Tech Stack
PythonAWSETLSQL
Job Description, Responsibilities & Requirements
Description language:
About the Position
Dataforest is seeking a Middle+ Data Engineer to join an exciting software development project in the field of Cybersecurity. Our client’s platform offers Identity Insights that drive real-time, multi-contextual digital forensics to protect new account opening workflows and expose fake accounts already in a customer database.
Responsibilities
- Maintain data processing architecture using Python and PySpark that interacts with ScyllaDB & PostgreSQL.
- Work with Databricks infrastructure to ensure reliability, scalability, and cost efficiency.
- Proactively identify bottlenecks and suggest technical improvements.
Requirements
- 3+ years of hands-on experience in data engineering;
- 3+ years of commercial experience with Python;
- Advanced experience with SQL DBs (optimizations, monitoring, etc.);
- Advanced experience with PySpark;
- Solid understanding of ETL principles (architecture/monitoring/alerting/search and resolve bottlenecks);
- Familiar with AWS infrastructure (boto3, S3 buckets, etc);
- Experience working with large volumes of data;
- Understanding the principles of medallion architecture.
Nice to Have
- Cassandra/Scylla
- PostgreSQL
- Pandas
- Experience with Kafka and Redis
We Offer
- Working in a fast-growing company;
- Great networking opportunities with international clients, challenging tasks;
- Personal and professional development opportunities;
- Competitive salary fixed in USD;
- Paid vacation and sick leaves;
- Flexible work schedule;
- Friendly working environment with minimal hierarchy;
- Team building activities, corporate events.
About the Company
Join Dataforest and be part of a dynamic team that is shaping the future of cybersecurity.