LOADING...
← Portfolio
Fastino logo

Fastino

1000x faster LLM inference.

Pre-seed2024AInew model architectures

Revolutionary inference approach that achieves OOM speedups and CPU compatibility. Early benchmarks show sub-millisecond response times. Inherently hallucination-resistant and optimal for sensitive enterprise workflows like structured outputs, PII masking, etc.

Co-investors
  • Microsoft M12
  • Insight Venture Partners