
Cerebras Systems 是一家前沿硬件公司,利用晶圆级技术设计并制造最大、最快的 AI 处理器。通过将内存与计算共置,其系统消除了昂贵的内存传输,从而实现快速的 AI 推理和开发。该项目对于突破 AI 硬件性能边界、推动新芯片范式具有重要意义。
新闻
you are not ready for cerebras inference ask. click. done. swe 1.7 is really fun to use btw. https://x.com/dabit3/status/2077483917263687807
对话 Cerebras CEO :手握 250 亿积压订单,AI 算力需求早就订满,我们不是「建了等客来」
i have a sneaky suspicion openai made a break-through in inference design, GPT 5.6 thinks for longer, consumes more tokens but at a much cheaper rate: more tokens-per-sec, cheaply served = more test-time compute that’s why it’s reportedly smashed any agentic benchmark. inference is the unlock. th
GPT-5.6 三兄弟,最快今天(7月7日)正式开放 • 上下文拉到150万token • Sol Ultra 跑在Cerebras上,生成速度约750 tokens/s——现在的10倍 定价更是刀刀见血: • Sol 输入$5/输出$30 • 对手Fable 5 是$10/$50 直接打一半价格 实测结果很微妙: • Sol vs Fable 5,五五开 • Sol赢3D建模,Fable 5赢游戏逻辑 但Fable 5安全限制太严,80%请求被降级分流回Opus 4.8 硬件这条线也在动: • Codex Micro,7月15日开售 • OpenAI 第一款硬件,是13个机械键+摇杆
GPT-5.6 三兄弟,最快今天(7月7日)正式开放 • 上下文拉到150万token • Sol Ultra 跑在Cerebras上,生成速度约750 tokens/s——现在的10倍 定价更是刀刀见血: • Sol 输入$5/输出$30 • 对手Fable 5 是$10/$50 直接打一半价格 实测结果很微妙: • Sol vs Fable 5,五五开 • Sol赢3D建模,Fable 5赢游戏逻辑 但Fable 5安全限制太严,80%请求被降级分流回Opus 4.8 硬件这条线也在动: • Codex Micro,7月15日开售 • OpenAI 第一款硬件,是13个机械键+摇杆
It is a 2 to 4T param model. They are serving it across 70-100 wafers. To get healthy serving characteristics, they are essentially putting at most one layer per wafer, and the model is in the ballpark of 70-90 layers. There's a couple of different ways this could be served and model sizes implied
据官方公告,币安将于 7 月 6 日 13:30(UTC)在 Binance Stocks 新增 10 个股票交易标的,包括 Cerebras Systems(CBRS)、Quantinuum(QNT)、Strategy 优先股 STRC、Antalpha(ANTA)及多只 ETF 等。https://wublock123.com/news/binance-stocks-adds-10-assets-including-strc-qnt-cbrs-64078
Most people should probably update their priors on the state of open-source speech-to-speech. It's honestly kind of mind-blowing. We teamed up with @cerebras to build a fully open-source realtime voice demo (models + code) to show what's possible today. Demo : huggingface.co/spaces/smolage… Blog
Cerebras Systems reports strong growth after OpenAI partnership
事实上,GPT-5.3-Codex-Spark 是 2 月 12 日上线的第一个真实产品,是 Cerebras 推理引擎的第一份实战成绩单。 当时的速度:10-15 倍于 GPU Codex-Spark 在 WSE-3 上跑出 1000+ tokens/秒,标准 GPU 跑同款模型约 65 tokens/秒,差距 15 倍。 不过缺陷也很明显:成本和良率。 等到市场的情绪和炒作结束,谁能最有性价比地提供解决方案才最重要。 一台 CS-3 系统报价约 $2-3M+。跑 frontier model 需要 20-30 台:$40-90M 的硬件成本,还不算功耗(23kW × 25 台 =