unsloth-buddy
Zero-friction LLM fine-tuning skill for Claude Code, Gemini CLI & any ACP agent. Unsloth on NVIDIA · TRL+MPS/MLX on Apple Silicon. Automates env setup, LoRA training (SFT, DPO, GRPO, vision), post-hoc GRPO log diagnostics, evaluation, and export end-to-end. Part of the Gaslamp AI platform.
파일 탐색기
최종 버전 다운로드 (.zip)- marketplace.json
- compare_sample_1.png
- compare_sample_2.png
- compare_sample_3.png
- gaslamp.md
- index.html
- README.md
- gaslamp.md
- index.html
- index_zh-Hans.html
- index_zh-Hant.html
- README.md
- README_zh-Hans.md
- README_zh-Hant.md
- tiny_dpo_pairs.jsonl
- tiny_grpo_prompts.jsonl
- tiny_sft_messages.jsonl
- lessons.md
- skills.md
- user.md
- eval_compare.log
- gaslamp.md
- progress_log.md
- project_brief.md
- __init__.py
- common.py
- filesystem.py
- lifecycle.py
- report.py
- static_contract.py
- trace.py
- lifecycle_smoke.csv
- negative_controls.csv
- skill_activation.csv
- unsloth_buddy_rubric.schema.json
- test_filesystem_lifecycle.py
- test_report.py
- test_static_contract.py
- test_trace.py
- EVAL.md
- README.md
- run_skill_evals.py
- unsloth_gaslamp.png
- add_reflect_hint.py
- colab_training.py
- demo_server.py
- detect_env.py
- detect_system.py
- gaslamp_callback.py
- init_project.py
- llamacpp.py
- mlx_eval_template.py
- mlx_eval_vision_template.py
- mlx_gaslamp_dashboard.py
- mps_grpo_example.py
- reflect.py
- search_design.py
- setup_colab.py
- terminal_dashboard.py
- unsloth_dpo_example.py
- unsloth_grpo_example.py
- unsloth_mlx_sft_example.py
- unsloth_mlx_vision_example.py
- unsloth_sft_example.py
- unsloth_vision_example.py
- data.md
- demo_builder.md
- interview.md
- chat_ui.html
- dashboard.html
- demo_llm_crisp.html
- demo_llm_dark.html
- demo_vlm_crisp.html
- demo_vlm_dark.html
- gaslamp.png
- gaslamp_template.md
- .gitignore
- AGENTS.md
- gemini-extension.json
- LICENSE
- README.md
- README_zh-Hans.md
- README_zh-Hant.md
- SKILL.md
// repository documentation
Was this content helpful?
(0 ratings)
