Inference

AI model training and hosting platform.

FOUNDED

2022

INDUSTRY

AI/ML

DevTools

TEAM SIZE

10

COMPANY WEBSITE

Request an intro →

About

Inference.net is a global AI inference platform designed to make deploying large language models (LLMs) fast, affordable, and developer-friendly. Blazing fast inference for open-source, custom, and fine-tuned AI models at massive scale. Automatically capture and optimize your LLM performance in production.

Fully compatible with the OpenAI API, Inference.net allows developers to switch by changing just a single line of code. Its infrastructure aggregates underutilized compute capacity across data centers, functioning as a spot market for perishable compute resources.

The platform supports a variety of use cases, including real-time chat, data extraction, and batch inference. Founded in 2023 and based in San Francisco, we were one of the first checks in Inference. They are now backed by top venture capital firms and industry experts. Inference was founded in 2022 by Sam Hogan and Abe (Ibrahim) Ahmed.

Want to join our portfolio?

If you have an idea you are excited about that fits our ethos, start an application. One of our team members will get back to you within a month.