The 940MX does run trf, at 1,158 words/s with batch 32 inside 2 GB, 6.2x its host CPU. The earlier claim that current PyTorch wheels cannot target sm_50 was only half right: Maxwell kernels were dropped from the cu128 and cu129 builds starting torch 2.8, which is what `pip install torch` now resolves to, but the cu126 build of 2.7.1 still ships sm_50 and works. .venv-trf-gpu pins torch==2.7.1+cu126 for this and stays separate from .venv, whose cupy runs on nvidia-* 12.9 wheels that torch would downgrade to 12.6. Drop the editorial sentence from the generated model card; the table states it. |
||
|---|---|---|
| .. | ||
| upstream | ||
| CONTRIBUTING-GUIDE.md | ||
| MODELS.md | ||