개발자
Tenstorrent 하드웨어에서 모델을 빠르게 실행하세요. 두 개의 오픈 소스 SDK로 가능한 한 금속에 가까이 접근하거나 AI 컴파일러에 맡길 수 있습니다.
모델
Explore models optimized on Tenstorrent hardware.
Don’t see yours listed? Check out TT-Forge, our compiler, to get other models running today.
모델48
Sep 29Sep 28Sep 28Sep 28Sep 16Sep 11Sep 11Sep 4Sep 4Sep 2
Fix fused scale-mask softmax tile-padding leakage at non-32 widths
easy
Fix u32-to-fp32 double rounding in typecast
medium
Improve div_no_nan accuracy to 1 ULP
medium
Improve BF16 reciprocal rounding on Wormhole
medium
SFPLOADMACRO produces wrong results under coverage instrumentation (Reciprocal fails 46/153 on Blackhole)
medium
ttnn.quantize/requantize uint8 lower-bound saturation
easy
Remove legacy sqrt/rsqrt/reciprocal compatibility paths
hard
Fix INT_MIN correctness in int32 div, remainder, fmod, and scalar promotion
hard
DRAM-sharded decode matmul: the in0 multicast is one hop per activation shard
hard
binary_ng/ternary: replicate the broadcast row with local NoC copies instead of ~2000 scalar stores per tile
medium