You
One class is 3% of the data and oversampling barely helps.
ChatGPT
Change the metric before changing the data. Accuracy is meaningless here; precision-recall AUC and a threshold chosen from the business cost of each error type usually reveal that the model was fine and the decision rule was wrong.