Hi, My Name is
I’m a technologist with 4+ years of experience in PL/SQL, ETL development, and scalable system design. Currently pursuing my Master’s in Computer Science at Texas A&M University-Kingsville, I’m passionate about solving real-world problems through data-driven and software engineering solutions.
Hi there! I’m Shantanu Fuke, a Data Engineer with 4+ years of experience in PL/SQL, ETL development, and building scalable data pipelines. I’m currently pursuing my Master’s in Computer Science at Texas A&M University-Kingsville, with a strong focus on cloud data platforms, big data technologies, and analytics solutions. I’m actively seeking full-time roles in Data Engineering and related domains.



Sep 2025 - Present
⦿ Designing and developing Oracle APEX applications to support business workflows and payment processing systems..
⦿ Writing and optimizing PL/SQL procedures and functions for backend data processing..
⦿ Integrating APEX applications with external services and REST APIs for seamless data exchange..


Jan 2024 - Dec 2024
⦿ Built modular Python pipelines to convert large FASTA files into model-ready format, reducing processing time by 60%.
⦿ Applied CNN, RNN, and LSTM models for DNA sequence classification, improving prediction accuracy by 35% over baseline.
⦿ Structured the pipeline for reusability with clear separation of data loading, transformation, and model training stages.
⦿ Developed and optimized C++ code to generate mutated DNA sequences, accelerating synthetic dataset creation by 70%.
⦿ Used Amazon S3 to centralize raw and processed genomic data, streamlining access for preprocessing and training.
⦿ Built a Databricks workflow triggered by DNA file uploads to AWS S3, using PySpark to read sequences and append predictions from a trained Keras model, saving results as a CSV file and reducing manual workload by 80%.


Oct 2021 - July 2023
⦿ Streamlined GCP to Oracle RMS data transfer using PL/SQL procedures, reducing stock ledger upload time by 40%.
⦿ Built PL/SQL integration to load retail invoices into SAP ERP, improving workflow efficiency and data accuracy by 60%.
⦿ Led development of an Oracle Retail Cloud ETL program to automate location list creation, reducing manual effort by 70%.
⦿ Developed BI Reports for invoices, stock ledger, and deals to enhance insights in the Oracle Retail Suite.
⦿ Replicated Oracle transformations in PySpark/Spark SQL for benchmarking and cloud-readiness testing.
⦿ Simulated PL/SQL batch jobs in Apache Airflow to improve orchestration visibility and recovery.
⦿ Rebuilt legacy Oracle BI logic in Snowflake using modular dbt models as part of a cloud migration initiative.


April 2019 - October 2021
⦿ Built PL/SQL ETL to migrate 80% of Coupa PO data to Oracle EBS, saving 20+ hrs/week via batch automation.
⦿ Automated user creation in Oracle EBS using PL/SQL and email integration, improving onboarding efficiency by 60%.
⦿ Enhanced Oracle Forms/Reports and applied performance tuning to support high-volume transactions.
⦿ Scheduled and monitored data load jobs (invoices, POs, users) using Control-M for timely execution.
⦿ Simulated Control-M workflows in Apache Airflow to assess orchestration, dependencies, and alerting for cloud readiness.
⦿ Prototyped Oracle flat-file ingestion using Azure Data Factory to simulate cloud-native orchestration.
I am proficient in Python and have used it extensively for backend development, data analysis, and machine learning projects. I am familiar with libraries like NumPy, Pandas, Matplotlib, and Scikit-learn.


2023 - 2025
GPA: 3.92


2015 - 2019
GPA: 3.74
