개발자
Tenstorrent 하드웨어에서 모델을 빠르게 실행하세요. 두 개의 오픈 소스 SDK로 가능한 한 금속에 가까이 접근하거나 AI 컴파일러에 맡길 수 있습니다.
모델
Explore models optimized on Tenstorrent hardware.
Don’t see yours listed? Check out TT-Forge, our compiler, to get other models running today.
모델59
Aug 15Aug 6Aug 5Aug 3Jul 28Jul 22Jul 21Jul 21Jul 15Jul 10
ttnn.softplus takes the linear branch at exactly input*beta == threshold (strict < where reference uses <=)
medium
Generalize multi_scale_deformable_attn to support D values that are multiples of 16
medium
Fix Blackhole destination-reuse synchronization
hard
Fix SFPU RNG correlation and improve FP32 uniform random quality
hard
Gemma-2 (2B / 9B) text model bring-up using TTNN APIs
hard
Fuse per-channel quantize/dequantize with a scalar zero-point
medium
Equal-Count Welford Reduction Optimisation
hard
[sdpa_decode] fp32_dest_acc_en=True silently produces wrong results for most shapes, and hangs at 96 query heads
medium
Performance/precision: atan/asin/acos (fp32)
hard
Optimise `deg2rad` / `rad2deg` (fp32/bf16)
medium