|
|
Zero-dependency inference engine for Apple Silicon M-series chips. Implements virtual memory block page tables (16 tokens/block) with dynamic NVMe SSD offloading, enabling 100K+ token context processing on 16GB Macs without memory thrashing or OOM crashes. Includes custom Metal Shading Language compute kernels (kernel_gemma4_rmsnorm, GEMV) and an ANSI terminal TUI telemetry visualizer.
National menu & restaurant intelligence platform indexing 500,000+ data entries for sub-millisecond geospatial retrieval using SQLite FTS5 full-text search and grounded local Ollama LLMs. Studio site (wizardrylabs.burhanmoin.com) achieves perfect 100/100 Google Lighthouse scores across Performance, SEO, Accessibility, and Best Practices.
Full-stack conversational e-commerce shopping assistant for Shopify storefronts. Features multimodal Vision & OCR image processing for instant customer photo product matching, async worker queues, and automated fallback LLM routing.
High-concurrency gaming platform backend handling 100,000+ concurrent users. Replaced
Native macOS dynamic notch overlay assistant bringing screenshot vision, persistent AI chat, media controls, and system actions into an always-available desktop interface.
🌐 Wizardry Labs Studio • 💼 Portfolio • 📧 Email Me • 💼 LinkedIn
"Building high-performance AI infrastructure from first principles."



