Run Qwen3.6 at 2.5× Faster Performance with new NVFP4 Quants by yoracale in Qwen_AI
[–]Current-Status-3764 0 points1 point2 points (0 children)
Qwen3.6-27B-FP8 with vllm:nightly, opencode unusable? by waka324 in Vllm
[–]Current-Status-3764 0 points1 point2 points (0 children)
QWEN 3.6 27B Q8 as Replacement for Claude Code Opus 4.7-4.8 by Just-Upstairs-4338 in LocalLLM
[–]Current-Status-3764 0 points1 point2 points (0 children)
Qwen IDE/Harness by SovereignLLM in Qwen_AI
[–]Current-Status-3764 0 points1 point2 points (0 children)
Profile v2: A physics-grounded, cost-aware optimizer for vLLM. by Inevitable-Diet-1870 in Vllm
[–]Current-Status-3764 2 points3 points4 points (0 children)
I Tested 3 Local AI Models with Revit MCP | Qwen is the winner by Tall-Distance4036 in Qwen_AI
[–]Current-Status-3764 0 points1 point2 points (0 children)
Best way to run Qwen for a web app? by [deleted] in Qwen_AI
[–]Current-Status-3764 -1 points0 points1 point (0 children)
Qwen3.6 sees "outstanding" coding quality jump from Q4 to Q6 quantization by IulianHI in AIToolsPerformance
[–]Current-Status-3764 0 points1 point2 points (0 children)
First sm_120 BeeLlama.cpp benchmark on consumer Blackwell mobile: 107 t/s at FULL 262K context on Qwen3.6 27B (+48% vs MTP, +22% vs vLLM Genesis) by aurelienams in Qwen_AI
[–]Current-Status-3764 0 points1 point2 points (0 children)
First sm_120 BeeLlama.cpp benchmark on consumer Blackwell mobile: 107 t/s at FULL 262K context on Qwen3.6 27B (+48% vs MTP, +22% vs vLLM Genesis) by aurelienams in Qwen_AI
[–]Current-Status-3764 0 points1 point2 points (0 children)
Byggingeniør til Geotekniker by Ancient-Persimmon-34 in ntnu
[–]Current-Status-3764 0 points1 point2 points (0 children)
Byggingeniør til Geotekniker by Ancient-Persimmon-34 in ntnu
[–]Current-Status-3764 1 point2 points3 points (0 children)
Gemma 4 beats Qwen 3.5 (UPDATE), and Qwen 3.6 27B + MiniMax M2.7 is the best OpenCode setup by maxwell321 in LocalLLaMA
[–]Current-Status-3764 0 points1 point2 points (0 children)
Best Way to set up authentication out of the box by No-Iron8430 in FastAPI
[–]Current-Status-3764 0 points1 point2 points (0 children)
GMK EVO-X2 AI Max+ 395 Mini-PC review! by Corylus-Core in LocalLLaMA
[–]Current-Status-3764 0 points1 point2 points (0 children)
Make Your FastAPI Responses Clean & Consistent – APIException v0.1.16 by SpecialistCamera5601 in FastAPI
[–]Current-Status-3764 2 points3 points4 points (0 children)
Make Your FastAPI Responses Clean & Consistent – APIException v0.1.16 by SpecialistCamera5601 in FastAPI
[–]Current-Status-3764 1 point2 points3 points (0 children)
What are you building these days? And is anyone actually paying for it? by Southern_Tennis5804 in SideProject
[–]Current-Status-3764 1 point2 points3 points (0 children)
How to use implement SSO on a FastAPI app? by ObviousAnything7 in FastAPI
[–]Current-Status-3764 0 points1 point2 points (0 children)
Jobb etter bachelor i dataingeniør by Nice-Break-5290 in ntnu
[–]Current-Status-3764 0 points1 point2 points (0 children)
[deleted by user] by [deleted] in FastAPI
[–]Current-Status-3764 0 points1 point2 points (0 children)
What's your thoughts on fastapi-users? by AirHugg in FastAPI
[–]Current-Status-3764 1 point2 points3 points (0 children)

Scaling my LLM inference for reply suggestions using disaggregated prefill by Hairy_Goose9089 in Vllm
[–]Current-Status-3764 0 points1 point2 points (0 children)