ยท
AI & ML interests
One-click deployment of Open-source LLMs, on managed and dedicated GPUs.
Recent Activity
Organizations
Qwen3.6-27B-FP8 on One RTX 6000 Ada: Fast TTFT, 314 tok/s Decode Generation [Benchmark]
Gemma-4 31B + vLLM on RTX 6000 PRO : A Real-Load Benchmark