← Portfolio

Fastino
1000x faster LLM inference.
Pre-seed2024AInew model architectures
Revolutionary inference approach that achieves OOM speedups and CPU compatibility. Early benchmarks show sub-millisecond response times. Inherently hallucination-resistant and optimal for sensitive enterprise workflows like structured outputs, PII masking, etc.
Co-investors
- Microsoft M12
- Insight Venture Partners