[P0] Threshold + Params for big model. Part 2. #361
- Threshold + Params for big model;
- Scripts to compute, re-check for existing;
-
New Models
-
gpt-oss-120b -
DeepSeek-R1-0528 -
gemma-3-27b-it -
Qwen3-30B-A3B-Instruct-2507 -
gpt-oss-20b -
Qwen3-235B -
Instruction to do it
- Inference Validation finetuning;
- Fine-tuning Qwen 235
💬 Comments (2)
The threshold-calculation task is completed for the models listed above (except GTP-OSS). They haven’t deployed it to the chain yet. They will most likely be deployed after the vLLM update
🔄 Auto-synced from Issue #361 every hour.
GPT-OSS can be implemented after the vLLM update. Right now, it is being handled by community contributors from the bounty program https://discord.com/channels/1336477374442770503/1425189436748206171/1446142256900997152