Mohit Acharya

Computers, AI, and all things cool.

about writing projects
  • Building a minimal vLLM implementation (GPU, LLM, Inference) (Aug 01, 2026)
  • Speeding up Generation: From Speculative Decoding to DSpark (Speculative decoder, LLM, Inference) (Jul 04, 2026)

© 2026 Mohit Acharya.