1. Home
  2. Companies
  3. FriendliAI
FriendliAI logoFR

FriendliAI

About

FriendliAI, founded in 2021, is a technology company spun out of Seoul National University research. It operates from Seoul, South Korea, and San Francisco, USA, focusing on AI/ML, enterprise software, and cloud infrastructure.

The company's core business is providing a fast AI inference platform for deploying and running large language models. Its technical work centers on AI inference optimization, GPU kernel development, caching, continuous batching, speculative decoding, and parallel inference. The FriendliAI Inference Platform claims to deliver over 2x faster inference and provides instant access to more than 590,000 models from Hugging Face.

FriendliAI offers enterprise-grade solutions including Dedicated Endpoints and container-based deployment options, both backed by a 99.99% uptime SLA. The company's stated mission is to make deploying large language models fast, affordable, and reliable.

Similar companies

Hyperbolic Labs logoHL

Hyperbolic Labs

Hyperbolic Labs operates an Open-Access AI Cloud providing a GPU marketplace and serverless AI inference service for developers.

1 job
fal logoFA

fal

fal is a generative media cloud platform providing developers with fast API access to over 600 production-ready AI models for generating images, video, audio, and 3D content.

LiteLLM logoLI

LiteLLM

LiteLLM is an open-source AI Gateway providing a unified OpenAI-compatible interface to access over 100 large language models from various providers.

CoreWeave logoCO

CoreWeave

CoreWeave is a specialized cloud platform providing GPU infrastructure for training and deploying large AI models, with data centers in the US and Europe.

UR

uRun

uRun provides the infrastructure layer for interactive generative AI, offering real-time, stateful inference at scale via its Inference Cloud Platform.

Bland logoBL

Bland

Bland builds an enterprise voice AI platform that automates phone calls with human-sounding AI agents, featuring low-latency real-time processing and native support for over 40 languages.