J-Zero
Official Implementation of "J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data" (https://arxiv.org/abs/2608.26582)
파일 탐색기
- overview.png
- __init__.py
- bleu_penalty.py
- config.py
- hint_delta.py
- local_backend.py
- local_render.py
- local_train.py
- lora_compat.py
- multi_round.py
- parse.py
- phase1.py
- phase1_sample.py
- phase2.py
- phase3.py
- prompts.py
- lora_4b.sh
- lora_8b.sh
- lora_llama3b.sh
- recreate_verl_env.sh
- .gitignore
- LOCAL.md
- README.md
- build_dummy.py
- gen_eval_responses.py
- llm_recheck.py
- run_solver_eval.sh
- serve_recheck_judge.sh
- launch_rzero_4b.sh
- launch_rzero_8b.sh
- launch_rzero_llama3b.sh
- evaluate.py
- run_parallel.sh
- upload_to_parquet.py
- question_generate.py
- run_parallel.sh
- answer_equiv.py
- challenger_reward.py
- rzero_batch_manager.py
- solver_reward.py
- start.sh
- start_vllm_server.py
- stop.sh
- challenger_train.sh
- cleanup_between_phases.sh
- env_prelude.sh
- README.md
- run_rzero.sh
- RUNS.md
- solver_train.sh
- aime24.parquet
- aime25.parquet
- amc23.parquet
- gsm8k.parquet
- math500.parquet
- minervamath.parquet
- olympiadbench.parquet
- eval_bbh.py
- eval_mmlupro.py
- eval_supergpqa.py
- prompts.jsonl
- reference_gpt4turbo.json
- configs.yaml
- qwen36_nothink.jinja
- serve_judge_vllm.sh
- collect_results.py
- gen_http.py
- gen_vllm.py
- prep_ifeval.py
- run_em_dp4.sh
- run_eval.sh
- score_ifeval.py
- to_alpaca_outputs.py
- api_config.yaml
- arena-hard-v2.0.yaml
- gen_answer_config.yaml
- gemini-2.0-flash-001.jsonl
- o3-mini-2025-01-31.jsonl
- question.jsonl
- add_markdown_info.py
- bedrock_utils.py
- completion.py
- judge_utils.py
- math_utils.py
- sglang_server.py
- gen_answer.py
- gen_judgment.py
- LICENSE
- README.md
- requirements-optional.txt
- requirements.txt
- show_result.py
- benchmark.py
- conversation.py
- elo.py
- elo_config_cw.py
- elo_helpers_cw.py
- matchup_selection_cw.py
- metrics.py
- scoring.py
- trueskill_solver_cw.py
- creative_writing_criteria.txt
- creative_writing_judging_prompt.txt
- creative_writing_prompts_v3.json
- human_writing_profile.json
- negative_criteria.txt
- pairwise_prompt.txt
- slop_list.json
- slop_list_bigrams.json
- slop_list_trigrams.json
- slop_phrase_prob_adjustments.json
- api.py
- file_io.py
- logging_setup.py
- .env_example
- .gitignore
- creative_bench_runs.zip
- creative_writing_bench.py
- generate_results_html.ipynb
- LICENSE
- model_name_subs.py
- README.md
- requirements.txt
- input_data.jsonl
- evaluation_lib.py
- instructions.py
- instructions_registry.py
- instructions_util.py
- LICENSE
- README.md
- requirements.txt
- alpaca_eval-pairwise_evaluator.patch
- README.md
- THIRD_PARTY.md
- launch_jzero_4gpu.sh
- run_jzero_4b.sh
- run_jzero_8b.sh
- run_jzero_llama3b.sh
- verl-core-rzero-fixes.patch
- vllm023-disable-trtllm-attn.patch
- build_dummy.py
- filter_rm.py
- gen_answers_rm.py
- run_parallel_rm.sh
- upload_to_parquet_rm.py
- question_generate.py
- run_parallel.sh
- challenger_reward.py
- challenger_reward_rm.py
- rzero_batch_manager.py
- solver_reward_rm.py
- build_pairs_challenger_neg.py
- build_pairs_subtask.py
- merge_pairs.py
- rm_io.py
- start.sh
- start_rm_server.py
- stop.sh
- start_answers.sh
- start_vllm_server_answers.py
- stop.sh
- challenger_train.sh
- cleanup_between_phases.sh
- env_prelude.sh
- rm_train.sh
- rm_train_loop.py
- run_jzero.sh
- solver_train.sh
- install_train_env_vllm023.sh
- requirements-eval-arena.txt
- requirements-eval-genbench.txt
- requirements-eval.txt
- requirements-train-vllm011.txt
- .env.example
- .gitignore
- LICENSE
- README.md
# CDN으로 사용하기
jsDelivrjsDelivr는 공개 GitHub 리포지토리를 별도 설정 없이 CDN으로 즉시 서빙합니다. 버전과 파일을 고르면 웹페이지에 바로 붙일 수 있는 링크와 예시 코드가 만들어집니다.
링크
예시
# 프로젝트 배지
// repository documentation
Was this content helpful?
(0 ratings)
