sixcat-eval
Six community LLM categories + one overall score. Fast local eval for OpenAI-compatible servers.
File Explorer
- check_release.py
- hermes_runner.py
- preflight.py
- status.py
- SKILL.md
- sixcat-readme-header-v0.2.0.png
- eval-run-findings.md
- harness-stdio.md
- answers.json
- answers_build.py
- prompts.json
- record_prompts.py
- replay_score.py
- run.json
- sheet.txt
- stdio_driver.py
- stdio_journal.jsonl
- stdio_prompts.json
- stdio_prompts_glm53.json
- stdio_run.json
- humaneval.jsonl
- ifeval.jsonl
- ifeval_100.jsonl
- tiny_arc.jsonl
- tiny_gsm8k.jsonl
- tiny_hellaswag.jsonl
- tiny_mmlu.jsonl
- tiny_truthfulqa.jsonl
- tiny_winogrande.jsonl
- __init__.py
- __main__.py
- client.py
- code.py
- context_preflight.py
- dataio.py
- instruct.py
- journal.py
- model-policies.json
- policy.py
- report.py
- run.py
- score.py
- selection.py
- tools.py
- GLM-5.3-Flash-chat-template-endthink.jinja
- __init__.py
- test_challenge_selection.py
- test_check_release.py
- test_code_safety.py
- test_concurrency.py
- test_context_preflight.py
- test_data_lint.py
- test_hermes_skill.py
- test_instruct.py
- test_journal.py
- test_letters.py
- test_loop_failures.py
- test_package_data.py
- test_parser_styles.py
- test_phase3_budgets.py
- test_phase4_both.py
- test_phase5_compare.py
- test_phase5_reporting.py
- test_policy.py
- test_retry_merge.py
- test_score.py
- test_self_auditing.py
- test_speed_receipts.py
- test_stdio_transport.py
- test_v05_protocol.py
- test_vendor_families.py
- test_vendor_truth_calibration.py
- __init__.py
- adjudicate_phase2.py
- calibrate_vendor_truth.py
- .gitignore
- LICENSE
- MANIFEST.in
- pyproject.toml
- README.md
- RELEASE_NOTES.md
- SELF_EVAL_REPORT.md
# Use via CDN
jsDelivrjsDelivr serves any public GitHub repository as a CDN with zero setup. Pick a version and a file to get a ready-to-paste link and snippet.
Link
Example
// repository documentation
Was this content helpful?
(0 ratings)
