Interesting. I did not know that. And still they got 15x speedup (https://www.cerebras.ai/blog/openai-gpt-oss-120b-runs-fastes...). I wonder how much additional speedup would be possible by really etching the model.
Interesting. I did not know that. And still they got 15x speedup (https://www.cerebras.ai/blog/openai-gpt-oss-120b-runs-fastes...). I wonder how much additional speedup would be possible by really etching the model.