Post @onusoz · /2026/08/16 · 05:41 AM View on Me: I want token subsidies Mom: We have token subsidies at home @victormustar · Aug 15, 2026 I deployed FREE public endpoint for Qwen3.8-27B no token needed, OpenAI-compatible, light rate limiting. Powered by Hugging Face Inference Endpoints (will be online for at least 72 hours). vision in, tool calls, 262K context, thinking dialable from xhigh (default) to off BF16 on one H200 using vLLM · real cost: $5/hr · 50 concurrent requests verified · ~60 tokens/seconds · 0.7s TTFT guide + code 👇 Show more
@onusoz · /2026/08/16 · 05:41 AM View on Me: I want token subsidies Mom: We have token subsidies at home @victormustar · Aug 15, 2026 I deployed FREE public endpoint for Qwen3.8-27B no token needed, OpenAI-compatible, light rate limiting. Powered by Hugging Face Inference Endpoints (will be online for at least 72 hours). vision in, tool calls, 262K context, thinking dialable from xhigh (default) to off BF16 on one H200 using vLLM · real cost: $5/hr · 50 concurrent requests verified · ~60 tokens/seconds · 0.7s TTFT guide + code 👇 Show more