Popular repositories Loading
-
agentic-coding-playbook
agentic-coding-playbook PublicSyv.ai's agentic coding playbook
-
Repositories
Showing 10 of 23 repositories
- HyperQwen Public
Serve large Qwen models fast on the GPUs you actually own. Qwen3.8-27B on a single 24 GB card with vLLM: 127 tok/s single-user (381 when the answer quotes the prompt), ~1,035 tok/s at 64 concurrent, 150k-262k context. vLLM patches, requant pipeline, benchmarks.
-
-
-
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Most used topics
Loading…