prometheus-eval
Evaluate your LLM's response with Prometheus and GPT4 π―
νμΌ νμκΈ°
μ΅μ’ λ²μ λ€μ΄λ‘λ (.zip)- alpha_optimal.png
- analogy.png
- finegrained_eval.png
- formats.png
- logo.png
- promixtheus.png
- terminology.png
- LLM_Evaluation_Recruiting_grounding.docx
- LLM_Evaluation_Recruiting_instruction.docx
- LLM_Evaluation_Recruiting_planning.docx
- LLM_Evaluation_Recruiting_reasoning_math.docx
- LLM_Evaluation_Recruiting_refinement.docx
- LLM_Evaluation_Recruiting_safety.docx
- LLM_Evaluation_Recruiting_theory.docx
- LLM_Evaluation_Recruiting_tool.docx
- LLM_Evaluation_Recruiting_grounding.qsf
- LLM_Evaluation_Recruiting_instruction.qsf
- LLM_Evaluation_Recruiting_planning.qsf
- LLM_Evaluation_Recruiting_reasoning_math.qsf
- LLM_Evaluation_Recruiting_refinement.qsf
- LLM_Evaluation_Recruiting_safety.qsf
- LLM_Evaluation_Recruiting_theory.qsf
- LLM_Evaluation_Recruiting_tool.qsf
- logo.png
- README.md
- README.md
- README.md
- any2str.py
- README.md
- README.md
- README.md
- any2str.py
- notes.txt
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- README.md
- basic_icl.txt
- inst_1k.txt
- inst_1k_v2.txt
- inst_1k_v3.help.txt
- inst_1k_v3.txt
- inst_1k_v4.help.txt
- inst_1k_v4.help.txt.md
- inst_1k_v4.txt
- inst_1k_v4.txt.md
- inst_1shot.txt
- inst_2k.txt
- inst_2k_v4.txt
- inst_help.txt
- inst_help_mini_yi.txt
- inst_help_v2 copy.txt
- inst_help_v2.txt
- inst_help_v2_yi.txt
- inst_help_v3-1k.txt
- inst_help_v3.txt
- inst_help_v4.txt
- inst_help_v5-1k.txt
- inst_help_v5-2k.txt
- inst_help_v6-1k.txt
- inst_help_v6-2k.txt
- inst_mini.txt
- inst_only.txt
- upload_hf.py
- .env.template
- .gitignore
- make_report.py
- poetry.lock
- pyproject.toml
- README.md
- requirements.txt
- run_api_inference.py
- run_base_inference.py
- run_chat_inference.py
- run_response_eval.py
- sample_evals.json
- sample_responses.json
- urial_conversation.py
- __init__.py
- pairwise_eval.py
- pairwise_example_output.jsonl
- pairwise_exchange_example_output.jsonl
- testdata_pairwise.jsonl
- utils_constants.py
- alpaca_eval.json
- autoj_pairwise.json
- feedback_collection_ood_test.json
- feedback_collection_test.json
- flask_eval.json
- hhh_alignment_eval.json
- mt_bench_eval.json
- mt_bench_human_judgement_eval.json
- preference_collection_ood_test.json
- vicuna_eval.json
- __init__.py
- data_loader.py
- __init__.py
- prometheus_utils.py
- vllm_utils.py
- __init__.py
- consistency.py
- get_report.py
- parser.py
- prompts.py
- README.md
- run_evaluate.py
- transitivity.py
- utils.py
- __init__.py
- judge.py
- litellm.py
- mock.py
- parser.py
- prompts.py
- utils.py
- vllm.py
- __init__.py
- test_judge.py
- test_parser.py
- Makefile
- poetry.lock
- pyproject.toml
- README.md
- test_absolute.py
- test_relative.py
- example_absolute.py
- example_relative.py
- build_documentation.yml
- build_pr_documentation.yml
- quality.yml
- tests.yml
- upload_pr_documentation.yml
- handbook.png
- introduction.mdx
- _toctree.yml
- deepspeed_zero3.yaml
- multi_gpu.yaml
- config_full_1.yaml
- config_full_10.yaml
- config_full_11.yaml
- config_full_12.yaml
- config_full_2.yaml
- config_full_3.yaml
- .gitignore
- merge_adapter.py
- prepare_dataset.py
- push_to_hub.py
- README.md
- run.sh
- train.sh
- launch.slurm
- conversation.py
- README.md
- run_dpo.py
- run_sft.py
- __init__.py
- configs.py
- conversation.py
- data.py
- model_utils.py
- release.py
- config_dpo_full.yaml
- config_sft_full.yaml
- __init__.py
- test_configs.py
- test_data.py
- test_model_utils.py
- .gitignore
- LICENSE
- Makefile
- README.md
- setup.cfg
- setup.py
- .gitignore
- LICENSE
- README.md
// repository documentation
Was this content helpful?
(0 ratings)
