最終更新:2025-08-19 (火) 02:32:44 (356d)
gpt-oss
Top / gpt-oss
https://openai.com/ja-JP/index/introducing-gpt-oss/
| モデル | レイヤー | パラメータ合計 | トークンあたりのアクティブパラメータ数 | エキスパート合計 | トークンあたりのアクティブエキスパート数 | コンテキスト長 | MXFP4 |
| gpt-oss-120b | 36 | 117b | 5.1b | 128 | 4 | 128k | 65GB |
| gpt-oss-20b | 24 | 21b | 3.6b | 32 | 4 | 128k | 14GB |
パラメータ数
Component 120b 20b MLP 114.71B 19.12B Attention 0.96B 0.64B Embed + Unembed 1.16B 1.16B Active Parameters 5.13B 3.61B Total Parameters 116.83B 20.91B Checkpoint Size 60.8GiB 12.8GiB
モデルカード
https://openai.com/ja-JP/index/gpt-oss-model-card/
https://cdn.openai.com/pdf/419b6906-9da6-406c-a19d-1bb078ac7637/oai_gpt-oss_model_card.pdf
gpt-oss-120b
- LM Studio
Apple M1 Ultra 40.22tk/s Apple M1 Max (64GB) 13.42tk/s Keep Model in Memoryをdisableに GeForce RTX 4060 Ti (16GB)+13900K+RAM 128GB GPUオフロード:6 Force Model Expert Weights onto CPU 5.1tk/s GPU利用率 (GPU-ZのGPU Load):10%、VRAM 1.3GB GeForce RTX 4060 Ti (16GB)+13900K+RAM 128GB GPUオフロード:0 5.3tk/s GPU利用率 (GPU-ZのGPU Load):10%、VRAM 1.1GB - o4-mini?
gpt-oss-20b
- LM Studio
GeForce RTX 3090 81.20tk/s Apple M1 Ultra 59.94tk/s GeForce RTX 4060 Ti (16GB) 71.83tk/s GPU利用率 (GPU-ZのGPU Load):90%、VRAM 11.7GB GeForce RTX 4060 Ti (16GB) + 13900K Force Model Expert Weights onto CPU 13tk/s GPU利用率 (GPU-ZのGPU Load):16%、VRAM 2.1GB GeForce RTX 4060 Ti (16GB) + 13900K GPUオフロード:0、KVキャッシュオフロード有効 7tk/s GPU利用率 (GPU-ZのGPU Load):11%、VRAM 1.2GB GeForce RTX 4060 Ti (16GB) + 13900K GPUオフロード:0、KVキャッシュオフロード無効 7tk/s GPU利用率 (GPU-ZのGPU Load):11%、VRAM 1.1GB Snapdragon X Plus (RAM 32GB) 28.12tk/s - o3‑mini?

