Unfortunately, the lack of an input cache discount makes it prohibitively expensive for most use cases that aren't one-shot prompts.
well same applies to GPT OSS 120. Qwen is just the much smarter model of the 2 public options on Cerebras.
well same applies to GPT OSS 120. Qwen is just the much smarter model of the 2 public options on Cerebras.