Developers
Get your models up and running fast on Tenstorrent hardware. With two open source SDKs, you can get as close to the metal as possible, or let our AI compiler do the work.
Models
Explore models optimized on Tenstorrent hardware.
Don’t see yours listed? Check out TT-Forge, our compiler, to get other models running today.
Models59
Aug 15Aug 6Aug 5Aug 3Jul 28Jul 22Jul 21Jul 21Jul 15Jul 10
ttnn.softplus takes the linear branch at exactly input*beta == threshold (strict < where reference uses <=)
medium
Generalize multi_scale_deformable_attn to support D values that are multiples of 16
medium
Fix Blackhole destination-reuse synchronization
hard
Fix SFPU RNG correlation and improve FP32 uniform random quality
hard
Gemma-2 (2B / 9B) text model bring-up using TTNN APIs
hard
Fuse per-channel quantize/dequantize with a scalar zero-point
medium
Equal-Count Welford Reduction Optimisation
hard
[sdpa_decode] fp32_dest_acc_en=True silently produces wrong results for most shapes, and hangs at 96 query heads
medium
Performance/precision: atan/asin/acos (fp32)
hard
Optimise `deg2rad` / `rad2deg` (fp32/bf16)
medium