Hardware teams struggle to implement LLMs efficiently on FPGAs despite the potential cost/performance benefits. A consultancy could provide architecture reviews, RTL optimization, and benchmarking.
Offer fixed-price optimization sprints for specific model architectures, plus ongoing support contracts.
Start with basic consulting for common open-source models before expanding to custom architectures.
The niche nature limits market size outside semiconductor companies.