×
Model Studio

What It Actually Takes to Run Qwen3.8-27B Locally

This article breaks down what it takes to run Qwen3.8-27B locally and offers a hosted alternative via Alibaba Cloud's Token Plan.

Model Studio Token Plan for Individual One Subscription for Every AI Model, Up to 3x More Value

The Model Studio Token Plan for Individual is now live — featuring the exclusive debut of Qwen3.8-Max.

Building an LLM Security Layer with Alibaba Cloud AI Gateway

A practical guide to blocking prompt injections, masking PII, and enforcing caller authentication at Alibaba Cloud AI Gateway

Qwen3.8-Max: A New Bar for Coding and Cowork

Today, we are officially releasing Qwen 3.8-Max, the most capable model in the Qwen family to date.

Alibaba Unveils Qwen3.8-Max: Its Largest and Most Capable Flagship Model to Date

Qwen3.8-Max exhibits advanced capabilities in coding, real-life work, research, and long-horizon tasks.

Alibaba Cloud AI Gateway로 구현하는 Self-Routing Multi-LLM 아키텍처

Alibaba Cloud AI Gateway와 Model Studio를 활용해 Self-Routing Multi-LLM 아키텍처를 구축하고, AI Fallback, 운영 모니터링, 비용 효율화를 구현하는 방법을 소개합니다.

Building a Self-Routing Multi-LLM Architecture with Alibaba Cloud AI Gateway

Learn how to build a Self-Routing Multi-LLM architecture with Alibaba Cloud AI Gateway and Model Studio, covering AI Fallback, operational monitoring, and cost optimization.

Qwen-Audio-3.0-TTS: More Multilingual, Easier to Direct

Our text-to-speech model, now across 16 languages.

Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

Following the release of Qwen3.6-Plus, we are sharing an early preview of our next proprietary model: Qwen3.6-Max-Preview.

How to Use Alibaba Cloud Model Studio for Generative AI Applications

This article provides a step-by-step guide to building production-ready generative AI applications using Alibaba Cloud Model Studio.

AI on Alibaba Cloud: Complete Guide to Machine Learning Services

This article introduces a comprehensive architectural guide for scaling machine learning workloads on Alibaba Cloud.

DeepSeek V4-Flash dalam Skala Besar: Panduan Deployment Berbasis Benchmark

Memilih cara men-deploy large language model di lingkungan produksi adalah salah satu keputusan paling konsekuensial — sekaligus paling membingungkan — yang dapat diambil sebuah tim AI.

大規模な DeepSeek V4-Flash の活用:ベンチマーク主導のデプロイガイド

本番環境で大規模言語モデルをどのようにデプロイするかは、AI チームが最も頭を悩ませる判断の一つです。

대규모 DeepSeek V4-Flash: 벤치마크 기반 배포 가이드

대규모 언어 모델을 프로덕션에 배포하는 방법을 선택하는 것은 AI 팀이 내릴 수 있는 가장 중요하면서 어려운 결정 중 하나입니다.

Alibaba Cloud Expands Global AI Infrastructure with New Data Centers in France, Japan, Malaysia, and Mexico

Alibaba Cloud has announced the launch of new availability zones across four key markets: France, Japan, Malaysia, and Mexico.

Alibaba Upgrades HappyOyster 1.0 with Enhanced Interactivity for Content Creation

Alibaba has released upgrades to HappyOyster 1.0 with richer environmental interactions, expanded player controls, and rewind-able storylines.

Alibaba Cloud Expands AI Infrastructure in Japan with Launch of Fifth Data Center and New Model Service Platform

Alibaba Cloud today announced the launch of its fifth data center in Japan.

Alibaba Cloud ECS 上で Telegram 連携機能付きの OpenClaw をデプロイする

クラウドで独自の AI コーディングエージェントを実行するためのステップバイステップガイド

Alibaba Cloud ECS에 Telegram 통합 기능을 갖춘 OpenClaw 배포

클라우드에서 자체 AI 코딩 에이전트를 실행하기 위한 단계별 가이드

Model Studio Architecture: A Deep Dive into Alibaba Cloud’s GenAI Application Platform

Running a generative AI application in production usually means stitching together a model server, a vector database, retrieval logic, a tool layer, a...