At TT-Deploy JP, Tenstorrent set new records on language and video models, launched TT-Ascalon S RISC-V CPU IP for agentic AI, and joins ai&’s sovereign heterogenous inference platform with Tenstorrent Galaxy™ superclusters, a general-purpose system that can drop in beside GPUs or stands alone.
August 6, 2026 -
Tenstorrent, the AI compute company led by CEO Jim Keller, today at TT-Deploy JP set new performance records across language and video, launched TT-Ascalon S RISC-V CPU IP for agentic AI, and detailed its largest deployment to date, a general-purpose, heterogenous AI build in Japan. Each rests on the same foundation: a single architecture that runs major AI workloads faster than GPUs and scales from a licensable core to a Tenstorrent Galaxy™ supercluster over standard Ethernet.
That makes Tenstorrent’s Networked AI architecture a different kind of solution – open, general-purpose, flexible for heterogeneous or stand alone deployments, and backed by AI experts – that can withstand the constant change in the AI industry.
Continuing to build on previous performance, Tenstorrent shared new LLM and video with lip-sync and audio benchmarks. On the latest models enterprises are deploying right now, Tenstorrent Galaxy Blackhole superclusters post:
Different model families, one architecture, with capacity that grows near-linearly as Galaxies are added. Tenstorrent’s performance enables enterprises to scale premium inference workloads efficiently.
Launching today at TT-Deploy JP, Tenstorrent announced TT-Ascalon S, a compute-dense RISC-V CPU for agentic AI. Agentic AI leans on the CPU in a new way, gated less by raw compute than by orchestration, I/O, and latency, and TT-Ascalon S is built for it...
Click here to read more