Providers /Phala
Inference provider

Phala

Models
20
Data residency
No answer on record
00

Our take

Editorial by LLMap · updated Sep 3, 2026

Phala is a privacy-first inference host that does not train on prompts and offers a zero data retention option. Its 28-model catalogue is small but includes the cheapest tracked offer in this sample.

Use this for privacy-sensitive workloads where zero retention and no prompt training are hard requirements. Pick it for budget inference on small models — the cheapest tracked offer in this sample sits here. Skip it if you need verified SOC 2, a confirmed EU endpoint, or predictable throughput across every model you might use.

Strengths
  • Strong privacy defaults: no prompt training and an optional zero retention setting.
  • Cheapest entry point in the sample catalogue.
  • Highest-throughput model in the sample runs more than nine times faster than the lowest.
Trade-offs
  • Compliance and corporate transparency documentation is thin: headquarters country undisclosed, no URL on file, SOC 2 and EU endpoint both unverified.
  • Extreme throughput inconsistency across the catalogue — five of twelve sample offers fall below 50 tokens per second.
  • One model throughput is unmeasured in our data.
01

Point your tools here

We have not recorded what a router needs for Phala yet — no base URL and no docs link on file. That is our gap, not a sign Phala has no API; its own documentation is the place to look until we close it.

02

Privacy & data handling

These answers cover requests sent to Phala through OpenRouter, as OpenRouter records them, last read 4 hours ago. For Phala's own API we hold no answer: check its terms. SOC 2 and the data processing agreement below are hand-checked, not part of that sync.

Trains on prompts
✗No
Logs prompts
✗No
✓Offered
No answer on record
No answer on record

Prompt retention: none.

03

Models & pricing

Something wrong on this page? Tell us