
Cerebras Systems builds specialized AI compute hardware, including the Wafer-Scale Engine. The company is mentioned in context of an IPO and AI infrastructure discussions.
News
you are not ready for cerebras inference ask. click. done. swe 1.7 is really fun to use btw. https://x.com/dabit3/status/2077483917263687807
对话 Cerebras CEO :手握 250 亿积压订单,AI 算力需求早就订满,我们不是「建了等客来」
i have a sneaky suspicion openai made a break-through in inference design, GPT 5.6 thinks for longer, consumes more tokens but at a much cheaper rate: more tokens-per-sec, cheaply served = more test-time compute that’s why it’s reportedly smashed any agentic benchmark. inference is the unlock. th
GPT-5.6 三兄弟,最快今天(7月7日)正式开放 • 上下文拉到150万token • Sol Ultra 跑在Cerebras上,生成速度约750 tokens/s——现在的10倍 定价更是刀刀见血: • Sol 输入$5/输出$30 • 对手Fable 5 是$10/$50 直接打一半价格 实测结果很微妙: • Sol vs Fable 5,五五开 • Sol赢3D建模,Fable 5赢游戏逻辑 但Fable 5安全限制太严,80%请求被降级分流回Opus 4.8 硬件这条线也在动: • Codex Micro,7月15日开售 • OpenAI 第一款硬件,是13个机械键+摇杆
GPT-5.6 三兄弟,最快今天(7月7日)正式开放 • 上下文拉到150万token • Sol Ultra 跑在Cerebras上,生成速度约750 tokens/s——现在的10倍 定价更是刀刀见血: • Sol 输入$5/输出$30 • 对手Fable 5 是$10/$50 直接打一半价格 实测结果很微妙: • Sol vs Fable 5,五五开 • Sol赢3D建模,Fable 5赢游戏逻辑 但Fable 5安全限制太严,80%请求被降级分流回Opus 4.8 硬件这条线也在动: • Codex Micro,7月15日开售 • OpenAI 第一款硬件,是13个机械键+摇杆
It is a 2 to 4T param model. They are serving it across 70-100 wafers. To get healthy serving characteristics, they are essentially putting at most one layer per wafer, and the model is in the ballpark of 70-90 layers. There's a couple of different ways this could be served and model sizes implied
据官方公告,币安将于 7 月 6 日 13:30(UTC)在 Binance Stocks 新增 10 个股票交易标的,包括 Cerebras Systems(CBRS)、Quantinuum(QNT)、Strategy 优先股 STRC、Antalpha(ANTA)及多只 ETF 等。https://wublock123.com/news/binance-stocks-adds-10-assets-including-strc-qnt-cbrs-64078
Most people should probably update their priors on the state of open-source speech-to-speech. It's honestly kind of mind-blowing. We teamed up with @cerebras to build a fully open-source realtime voice demo (models + code) to show what's possible today. Demo : huggingface.co/spaces/smolage… Blog
Cerebras Systems reports strong growth after OpenAI partnership
事实上,GPT-5.3-Codex-Spark 是 2 月 12 日上线的第一个真实产品,是 Cerebras 推理引擎的第一份实战成绩单。 当时的速度:10-15 倍于 GPU Codex-Spark 在 WSE-3 上跑出 1000+ tokens/秒,标准 GPU 跑同款模型约 65 tokens/秒,差距 15 倍。 不过缺陷也很明显:成本和良率。 等到市场的情绪和炒作结束,谁能最有性价比地提供解决方案才最重要。 一台 CS-3 系统报价约 $2-3M+。跑 frontier model 需要 20-30 台:$40-90M 的硬件成本,还不算功耗(23kW × 25 台 =