1. Home
  2. Companies
  3. FriendliAI
FriendliAI logoFR

FriendliAI

About

FriendliAI, founded in Seoul in 2021, specialises in optimising how large language models are served. The company's core technology accelerates AI inference - the process of running trained models in production - through techniques including continuous batching, speculative decoding, custom GPU kernels, and parallel inference. Its platform promises more than twice the inference speed of standard approaches and provides instant access to over 590,000 models from Hugging Face.

The company offers three main products: a managed AI inference platform, dedicated endpoints for enterprise clients running custom or fine-tuned models, and container-based deployment solutions. Enterprise offerings come with 99.99% uptime SLAs. The firm serves customers across AI research, telecommunications, and the broader technology sector.

FriendliAI opened an office in San Francisco in 2026 alongside its original Seoul headquarters. Financial terms and funding details have not been publicly disclosed.

Job at FriendliAI

Explore 1 job at FriendliAI and find your next opportunity.

FriendliAI logoFR

Software Engineer – Platform Security

FriendliAI

San Francisco, California, United States (Hybrid)

5 days ago

Similar companies

Oumi logoOU

Oumi

Oumi is an open-source AI platform that enables developers and enterprises to build, evaluate, and deploy custom AI models across the entire model lifecycle.

1 job
LILT logoLI

LILT

LILT is an AI-powered translation and localization platform that combines neural machine translation and NLP with a network of specialized linguists to deliver enterprise translation services across 100+ languages.

1 job
ElastixAI logoEL

ElastixAI

ElastixAI is a startup building scalable, energy-efficient AI inference solutions using FPGAs and full-stack optimization as an alternative to GPU infrastructure.

1 job
Featherless AI logoFA

Featherless AI

Featherless AI provides a serverless inference platform for instant access to over 20,000 open-source AI models, featuring rapid hot-swapping and significant cost reduction.

d-Matrix logoD-

d-Matrix

d-Matrix builds computing platforms using digital in-memory compute technology, designed specifically for efficient and scalable AI inference workloads.

Cohere logoCO

Cohere

Cohere builds secure, private foundational AI models for enterprise deployment, with research teams in Toronto and worldwide.