bumblecore
An LLM training framework built from the ground up, featuring a custom BumbleBee architecture and end-to-end support for multiple open-source models across Pretraining → SFT → RLHF/DPO.
File Explorer
Download Latest Version (.zip)- bug_report.md
- custom.md
- feature_request.md
- CODE_OF_CONDUCT.md
- CONTRIBUTING.md
- PULL_REQUEST_TEMPLATE.md
- SECURITY.md
- dpo_training_loss.png
- dpo_training_rewards_accuracies.png
- pretrain_training_loss.png
- sft_training_loss.png
- bumblechat.png
- bumblecore.jpg
- ds_z0_config.json
- ds_z1_config.json
- ds_z2_config.json
- ds_z3_config.json
- dpo_full.yaml
- dpo_lora.yaml
- bumblechat.yaml
- chat.yaml
- pretrain_full.yaml
- sft_full.yaml
- sft_lora.yaml
- alpaca_zh_demo.json
- dpo_messages_zh_demo.json
- dpo_zh_demo.json
- duolun_train.json
- glaive_toolcall_zh_demo.json
- messages_zh_demo.json
- pretrain_zh_demo.json
- CONFIG.md
- CONFIG_zh.md
- DATA_FORMAT.md
- DATA_FORMAT_zh.md
- FEATURES.md
- TUTORIAL.md
- TUTORIAL_zh.md
- styles.css
- black.png
- bumblebee.jpg
- logo.jpg
- white.png
- bumblechat.html
- config.json
- generation_config.json
- tokenizer.json
- tokenizer_config.json
- bumblechat.sh
- chat.sh
- dpo_full.sh
- dpo_lora.sh
- pretrain.sh
- sft_full.sh
- sft_lora.sh
- __init__.py
- modeling_bumblebee.py
- __init__.py
- arg_parser.py
- __init__.py
- train_config.py
- __init__.py
- data_formatter.py
- datasets.py
- preprocess.py
- __init__.py
- api.py
- inference.py
- streaming.py
- __init__.py
- loss.py
- __init__.py
- base_trainer.py
- dpo_trainer.py
- launcher.py
- pretrain_trainer.py
- sft_trainer.py
- __init__.py
- logger.py
- visualize_loss.py
- __init__.py
- api.py
- inference.py
- train.py
- test_arg_parser.py
- test_data_formatter.py
- test_datasets.py
- test_launcher.py
- conftest.py
- run_test.sh
- merge_lora.py
- run_merge_lora.sh
- .gitignore
- LICENSE
- pyproject.toml
- README.md
- README_zh.md
// repository documentation
Was this content helpful?
(0 ratings)
