AlphaSense announced that its Generative Search now runs on Cerebras’s high‑speed inference chips. The partnership cuts the time to first token by 88%, from 19.5 seconds to 2.3 seconds, letting the system deliver initial answers in as little as 10‑20 seconds. Faster model calls let AlphaSense keep the multi‑step research workflow—planning, routing, snippet evaluation—while still feeling interactive.
Why it matters
Analysts can get evidence‑based answers within minutes instead of hours, letting them iterate faster on earnings calls and market research.