
VCN #47: Fast Local
Join us for a hands-on build night where bringing a laptop is mandatory! π»
A local model that's too slow becomes dead weight, and tonight we aim to make it fast. π In VCN #47: Fast Local, weβll enhance the performance of your local rig to ensure it can operate efficiently.
Format:
- The Walkthrough: Discover where the time goes! Weβll cover GGUF quantization tradeoffs, llama.cpp vs vLLM, tensor parallelism, speculative decoding, KV-cache tuning, and how to measure tokens/sec accurately.
- The Benchmark Bar: We start by running Claude Code powered by z.ai and measuring the hosted speed. Itβs essential to know your target for optimizing your local endpoint. π
- Make it Fast: During the hands-on hour, quantize a model, create a fast serving endpoint using Nebius Token Factory credits, tune parameters, and watch your tokens/sec climb! β±οΈ
- Demos: Point your coding agent at your fresh fast endpoint and experience the improvements live.
By the end of the night, youβll leave with a quick local inference endpoint your coding agent can thrive on. π
Builders Only:
Bring your local rig from Bare Metal or just a model you wish to serve quickly.
π Doors Open: 7 PM
π Walkthrough Starts: 7:30 PM
π’ Location: Frontier Tower, Floor 10
Hosted by: Vibe Coding Nights β Rayyan Zahid (Immersive Commons), Michalis Vasileiadis (Hacker Bob), Eric Mockler (AI Geneticist), Devinder Sodhi (Learning Layer Labs).
Facilitator: Rayyan Zahid. Guest speaker TBD (open call).
π RSVP if your local model is private but painfully slow and you want it quick.
Frontier Tower members: Your ticket is complimentary! Please reach out to the team directly for a free RSVP. ποΈ
Frontier Tower π§βπ, 995 Market St, San Francisco, CA 94103, USA
Get directionsScan med kameraet β begivenheden Γ₯bnes i Somo-appen.









