Inference
AI model training and hosting platform.
FOUNDERS
FOUNDED
2022
INDUSTRY
AI/ML
DevTools
TEAM SIZE
10
COMPANY WEBSITE
Request an intro →
About
Inference.net is a global AI inference platform designed to make deploying large language models (LLMs) fast, affordable, and developer-friendly. Blazing fast inference for open-source, custom, and fine-tuned AI models at massive scale. Automatically capture and optimize your LLM performance in production.
Fully compatible with the OpenAI API, Inference.net allows developers to switch by changing just a single line of code. Its infrastructure aggregates underutilized compute capacity across data centers, functioning as a spot market for perishable compute resources.
The platform supports a variety of use cases, including real-time chat, data extraction, and batch inference. Founded in 2023 and based in San Francisco, we were one of the first checks in Inference. They are now backed by top venture capital firms and industry experts. Inference was founded in 2022 by Sam Hogan and Abe (Ibrahim) Ahmed.
Want to join our portfolio?
If you have an idea you are excited about that fits our ethos, start an application. One of our team members will get back to you within a month.


