OpenAI's Jalapeño Chip Outperforms Nvidia in Inference Benchmarks

· AI Analysis · AIssential

What happened

Chinese AI lab Z.ai has confirmed its GLM-5.3-Flash model is the anonymous 'Ox Alpha' that topped OpenRouter's rankings, doubling DeepSeek's performance. This model, released with open weights, is priced at a discounted $0.045 per task, approximately one-tenth of comparable rivals, and achieved its record performance on domestic chips.

Why it matters

The launch of Z.ai's GLM-5.3-Flash model, offering high performance at a tenth of the cost and running on domestic chips, signals a significant shift in AI accessibility and hardware independence, compelling practitioners to evaluate new, cost-effective models for their deployments.

Topics

Articles in this trend

Open in AIssential →