AI inference startup Wafer completes a $40 million Series A funding round with a valuation exceeding $200 million.

icon MarsBit
Share
AI summary iconSummary
AI and crypto news outlet MarsBit reports that AI inference startup Wafer has raised $40 million in a Series A round, pushing its valuation above $200 million. With only eight employees, the company achieved $8 million in annual recurring revenue (ARR) within three months. Wafer uses AI agents to optimize existing chips for faster, more cost-effective model execution on hardware such as NVIDIA and AMD. Tests showed that GLM 5.2 running on AMD MI355X delivered 80% of the throughput of NVIDIA B200 at half the cost. Clients include Vercel and Inworld AI. Vercel confirmed that Wafer’s GLM 5.2 throughput was double that of other serverless providers. Recent funding news indicates the company is now hiring across tech, growth, CEO office, and go-to-market roles in San Francisco.

Dynamic Beating AI News: Wafer, an AI inference startup with only eight team members, has raised $40 million in Series A funding, valuing the company at over $200 million. Previously, the company turned down acquisition offers from multiple cloud providers and inference service vendors. Wafer achieved $8 million in annual recurring revenue (ARR) from its inference business in just three months. It does not manufacture chips; instead, it helps models run faster and more cost-effectively on existing hardware. Wafer enables AI agents to automatically optimize models, inference engines, kernels, caching, quantization, and scheduling, then identifies the optimal configurations for different hardware such as NVIDIA and AMD. Previously, much of this work required manual tuning by inference performance engineers—Wafer aims to automate this process with AI. In its own tests, GLM 5.2 running on AMD MI355X achieved approximately 80% of the throughput of NVIDIA B200 at less than half the cost. Current customers include Vercel and Inworld AI. Vercel has also independently published test results showing that Wafer delivers roughly twice the throughput of other serverless providers when running GLM 5.2. Immediately after closing the funding round, Wafer began expanding its team, currently hiring for four roles: engineering, growth, CEO’s office, and GTM—all requiring full-time on-site presence in San Francisco five days per week. Engineering positions offer a base salary of $250,000 plus equity.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.