Lin Qiao
Showing 1–2 of 2 transcripts.
- Sequoia Capital10 min
How to Scale AI Application Inference 100x ft. Fireworks’ Lin Qiao
Fireworks addresses the critical gap between generic model capabilities and specific application needs by co-optimizing inference across quality, speed, and concurrency for over 100,000 variables. This strategic approach reduces inference costs by 10x to 100x through virtual cloud infrastructure that dynamically selects hardware and aligns data distributions, enabling rapid enterprise scaling. The platform has already demonstrated success in transforming high-cost operations into sustainable models, exemplified by a food chain expanding to 1,000 shops and a software firm serving 25 million developers within three months.
- Sequoia Capital39 min
Fireworks Founder Lin Qiao on the Power of Small Models to Democratize AI Use Cases
Lin Qiao, Sonya Huang, Pat Grady, Lynn Tiao
Founded in 2022 by former Meta PyTorch leaders Lynn Diao and others, Fireworks is a SaaS platform dedicated to compressing AI model deployment timelines from years to days through a specialized, PyTorch-native infrastructure. The company automates complex optimization tasks like quantization and semantic caching using handwritten CUDA kernels to enable high-performance, low-latency inference for small model stacks and fine-tuned enterprise workloads. By targeting the gap between research and production, Fireworks facilitates the migration of startups and traditional enterprises away from generic experimentation toward scalable, cost-efficient custom models that compete with larger monolithic systems.