Towards Effective and Efficient Continual Pre-training of Large Language Models | Read Paper on Bytez