| user: | anotherCodder |
| created: | September 6, 2024 |
| karma: | 23 |
| about: | Solo engineer. I build memra, an LLM inference engine in Rust and CUDA for RTX Blackwell: https://github.com/avifenesh/memra I also run a hosted inference API on it: https://inference.tiyuvta.ai |
| 1. | |
| 2. | Show HN: Qwen3.8-27B API, 140 tok/s on one GPU(inference.tiyuvta.ai) |
| 3. | |
| 4. | |
| 5. | |
| 6. | |
| 7. | |
| 8. | |
| 9. | |
| 10. | |
| 11. | |
| 12. | |
| 13. | |
| 14. | |
| 15. |