ling 3.0 flash/tiny base models

huggingface.co/inclusionAI/Ling-3.0-tiny-base huggingface.co/inclusionAI/Ling-3.0-flash-base-midtrain huggingface.co/inclusionAI/Ling-3.0-flash-base-30T huggingface.co/inclusionAI/Ling-3.0-tiny-base-midtrain huggingface.co/inclusionAI/Ling-3.0-tiny-base-30T These checkpoints correspond to different stages of the training process: Pretrained checkpoint have completed large-scale pretraining but have not undergone mid-training, WSM merging (or learning-rate decay), or post-training. Mid-trained checkpoint have completed mid-training but have not undergone WSM merging (or learning-rate decay) or post-training. Merged checkpoints have undergone WSM merging (or learning-rate decay) based on the mid-training checkpoints but have not undergone post-training. These checkpoints are released to support continued pretraining, fine-tuning, and further research. For the post-trained model, please see and see Ling-3.0-tiny and Ling-3.0-flash .

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论