- Measure Qwen3-Coder-30B-A3B, Qwen2.5-Coder-32B, and Tiel-Coder-35B-A3B at 64k and 32k context - Verify long-context degradation profile and MoE attention scaling - Add benchmark suite and results to BENCHMARK_MOE_CANDIDATES.md and benchmark_moe_results.json |
||
|---|---|---|
| .. | ||
| __init__.py | ||
| BENCHMARK_MOE_CANDIDATES.md | ||
| benchmark_moe_results.json | ||
| benchmark_suite.py | ||
| fast_downloader.py | ||
| measure_exact_moe_vram.py | ||
| run_moe_benchmark.py | ||