A new paper proposes integrating enlarged pretraining with structured pruning for efficient LLMs. It studies whether pretraining an oversized model before pruning beats training the target size from scratch. The pipeline aims to cut token costs while keeping deployable models within inference budgets.
Opening Kapyn…