Phala Confidential AI: 202.74 billion tokens over the past 24 hours. Top models: GLM-5.3 Flash 42.98% GLM-5.3 24.90% deepseek-v4.1-flash 13.25% https://lnkd.in/gEVMa8Fn TEE-backed AI infrastructure.
About us
Your open-source, trustless cloud. Powered by TEE, Governed by code, and Owned by you
- Website
-
https://phala.com
External link for Phala
- Industry
- Data Security Software Products
- Company size
- 11-50 employees
- Headquarters
- Newark, CA
- Type
- Privately Held
- Founded
- 2019
- Specialties
- Cloud, Web3, TEE, Computing, Trustless, AI, LLM, AGI, Security, and Zero Trust
Employees at Phala
Locations
-
Primary
Get directions
39899 Balentine Drive, Suite 200, Newark
Unit 202
Newark, CA 94560, US
-
Get directions
Singapore, SG
-
Get directions
39899 Balentine Drive, Suite 200, Newark
Unit 202
Newark, CA 94560, US
Updates
-
Privacy claims in AI should be verifiable at the workload level. SayGM moved its operator supply layer to Phala Cloud as confidential workloads. Each operator runs an approved image pinned to a digest inside its own hardware-attested confidential VM. Before a workload can serve traffic, SayGM’s registry checks its image hash. The attestation server and data-plane proxy run together inside the same measured workload, so the proof covers the code that receives the request and chooses its route. Operators bring their own upstream provider credentials. Those credentials are injected into the encrypted workload at deployment and remain sealed from SayGM and the host while the workload runs. The stack is built on dstack, the Apache 2.0 open-source TEE framework hosted by the Linux Foundation: https://lnkd.in/gaxjDBMy Live today: • 60 distinct coldkeys deployed • 18 operators active • 57 confidential VMs running SayGM shows how independent inference operators can serve traffic under one approved, verifiable runtime. Read the full case study: https://lnkd.in/dcrEzghA #ConfidentialAI #ConfidentialComputing #AIInfrastructure
-
-
Phala Confidential AI: 173.64 billion tokens over the past 24 hours. Top models: GLM-5.3 51.89% GLM-5.3 Flash 25.07% deepseek-v4.1-flash 10.61% https://lnkd.in/gTXc-TB7 TEE-backed AI infrastructure.
-
-
Phala Confidential AI: 254.51 billion tokens over the past 24 hours. Top models: GLM-5.3 70.49% Kimi K3 10.73% GLM-5.3 Flash 8.18% https://lnkd.in/gTXc-TB7 TEE-backed AI infrastructure.
-
-
Phala Confidential AI: 178.40 billion tokens over the past 24 hours. Top models: GLM-5.3 47.87% deepseek-v4-pro-0813 12.79% GLM-5.3 Flash 10.08% https://lnkd.in/gTXc-TB7 TEE-backed AI infrastructure.
-
-
Phala Confidential AI: 168.69 billion tokens over the past 24 hours. Top models: GLM-5.3 42.87% deepseek-v4.1-flash 19.24% GLM-5.3 Flash 12.00% https://lnkd.in/gTXc-TB7 TEE-backed AI infrastructure.
-
-
Phala Confidential AI: 55.18 billion tokens over the past 24 hours. Top models: GLM-5.3 Flash 49.10% GLM-5.3 14.71% Kimi K3 12.62% https://lnkd.in/gEVMa8Fn TEE-backed AI infrastructure.
-
-
Phala Confidential AI: 78.09 billion tokens over the past 24 hours. Top models: GLM-5.3 43.54% GLM-5.3 Flash 26.96% DeepSeek V4 Flash 11.46% https://lnkd.in/gTXc-TB7 TEE-backed AI infrastructure.
-
-
Qwen3 8B is live on Phala—served inside a TDX-attested GPU TEE. Private, verifiable inference for coding, long-horizon agents, reasoning. https://lnkd.in/gkUv9uSu 41K context. 8.2B parameter causal language model from the Qwen3 series. Reasoning, tool use, structured outputs. $0.11/M input · $0.45/M output. https://lnkd.in/e4yu_p_C
-
-
Phala Confidential AI: 50.72 billion billed input + output tokens over the past 24 hours. Top models: 1) z-ai/glm-5.3-flash 54.28% 2) deepseek/deepseek-v4-flash 17.77% 3) z-ai/glm-5.3 7.01% https://lnkd.in/gEVMa8Fn TEE-backed AI infrastructure.
-