Developers
Get your models up and running fast on Tenstorrent hardware. With two open source SDKs, you can get as close to the metal as possible, or let our AI compiler do the work.
Models
Explore models optimized on Tenstorrent hardware.
Don’t see yours listed? Check out TT-Forge, our compiler, to get other models running today.
Models48
Sep 29Sep 28Sep 28Sep 28Sep 16Sep 11Sep 11Sep 4Sep 4Sep 2
Fix fused scale-mask softmax tile-padding leakage at non-32 widths
easy
Fix u32-to-fp32 double rounding in typecast
medium
Improve div_no_nan accuracy to 1 ULP
medium
Improve BF16 reciprocal rounding on Wormhole
medium
SFPLOADMACRO produces wrong results under coverage instrumentation (Reciprocal fails 46/153 on Blackhole)
medium
ttnn.quantize/requantize uint8 lower-bound saturation
easy
Remove legacy sqrt/rsqrt/reciprocal compatibility paths
hard
Fix INT_MIN correctness in int32 div, remainder, fmod, and scalar promotion
hard
DRAM-sharded decode matmul: the in0 multicast is one hop per activation shard
hard
binary_ng/ternary: replicate the broadcast row with local NoC copies instead of ~2000 scalar stores per tile
medium