• File

Personal information hidden

This job seeker decided to hide his personal information and contact info, but you can send a message to him or suggest a job to him.

This job seeker has chosen to hide his personal information and contact info. You can contact him using this page: https://www.work.ua/resumes/18832146/

Junior big data engineer

Considering positions:
Junior big data engineer, Junior data engineer
City of residence:
Kyiv
Ready to work:
Kyiv, Remote

Contact information

This job seeker has hidden his personal information, but you can send him a message or suggest a job to him if you open his contact info.

Name, contacts and photo are only available to registered employers. To access the candidates' personal information, log in as an employer or sign up.

Uploaded file

Quick view version

This resume is posted as a file. The quick view option may be worse than the original resume.

Yevheniia Kaspruk
Junior Big Data Engineer
Email: [open contact info](look above in the "contact info" section)
LinkedIn: [open contact info](look above in the "contact info" section)
Telegram: @eugenykaspruk

Professional Summary
Aspiring Junior Big Data Engineer with a solid foundation in Data Science from American University Kyiv and hands-on data engineering experience.
Optimizing distributed data architectures, automating workflows, and constructing end-to-end ETL/ELT pipelines.
Possesses a strong analytical mindset and an exceptional attention to detail, combined with hands-on experience across the Apache ecosystem, modern
NoSQL databases, and AWS cloud infrastructure.

Technical Skills
Languages: Python, MySQL, PostgreSQL
Distributed Compute: Apache Spark (PySpark, Spark Streaming), Trino
Orchestration: Apache Airflow, Apache Kafka
Storage: AWS S3, MongoDB, Cassandra, HDFS, YARN
Cloud & DevOps: AWS (Certified Cloud Practitioner certification in progress), Docker, Git/GitHub

Professional Experience
Big Data Engineer Intern @ Grid Dynamics
12/2025 – 05/2026
Infrastructure: Built a distributed Big Data sandbox to test and validate integration across the Apache ecosystem.
ETL Pipelines: Developed modular Python and PySpark pipelines to automate large-scale data processing between HDFS, AWS S3, and NoSQL
databases.
Orchestration: Designed Apache Airflow DAGs to coordinate complex data processing workflows.
Storage: Configured and optimized HDFS, AWS S3, Cassandra, and MongoDB clusters for data staging.
Query Optimization: Tuned SQL queries to boost performance across distributed Trino engines.
Resource Management: Managed cluster resource allocation, queues, and job scheduling using YARN to maximize pipeline efficiency.
Architecture: Evaluated and selected the most efficient tools for analytics tasks.

Education
American University Kyiv
2024 – 2027: Bachelor of Science, Data Science

Certificates
📜 Certificates (link)
Languages
Ukrainian: Native Speaker
English: C1 / Full Professional
Spanish: B1 / Limited Working

Core Strengths
Attention to Detail
Problem-Solving
Fast Learner
Cross-Team Collaboration
Adaptability

Yevheniia Kaspruk 1

More resumes of this candidate

Similar candidates

All similar candidates


Compare your requirements and salary with other companies' jobs: