Developers working with AI models on Apple silicon need optimized inference engines.
A local inference engine tailored for Apple hardware could offer better performance and efficiency.
Developers and AI startups would pay for improved performance and ease of integration.
The first version could focus on supporting a few popular models.
The biggest risk is competition from larger AI frameworks.