Microsoft Phi-3.5 MoE Instruct

Metin & SohbetMevcut Değil
Microsoft tarafındanModel Kimliği: phi-3-5-moe-instruct

Mixture-of-experts Phi-3.5: 42B total / 6.6B active params. 128k context, multilingual.

Durum
Mevcut Değil
Bağlam
131.072 token
Maks. Çıkış
4.096 token
Giriş → Çıkış
Metin → Metin
Geliştirici
Microsoft
Güncellendi
25 Haziran 2026

Microsoft Phi-3.5 MoE Instruct şu anda kullanılamıyor

Şu anda kullanılamıyor: bu model devre dışı bırakılmıştır.

Bu sayfadaki ayrıntıları yine de okuyabilirsiniz. Hemen karşılaştırılabilir bir modeli çalıştırmak için aşağıdaki mevcut alternatiflerden birini seçin.

Alternatiflere Git
01

Karşılaştırılabilir modeller

Bu kategorideki tümü
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    $12,00/1M giriş

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    $6,00/1M giriş

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    $4,80/1M giriş

02

Playground

Microsoft Phi-3.5 MoE Instruct'ı deneyin

Sohbet

Şu anda kullanılamıyor

Şu anda kullanılamıyor: bu model devre dışı bırakılmıştır.

Oyun alanı devre dışıdır. Aynı kategoride karşılaştırılabilir modeller bulabilirsiniz: Alternatifleri görüntüle

Microsoft Phi-3.5 MoE Instruct'ı deneyin

Bir mesaj gönderin. Yanıt model işini bitirdikten sonra tamamen gelir (akışsız).

Sistem istemi
Maks. yanıt uzunluğu (token)

Bu çalıştırma

Fiyat yok – şu anda kullanılamıyor.

Yeni misiniz?

Google ile kaydolduğunuzda 5 ücretsiz kredi ($0,05)

Kaydolduktan 24 saat sonra kullanılabilir, günde 5 çalıştırmaya kadar ve çalıştırma başına en fazla 2 kredi. Diğer oturum açma yöntemleri kredi olmadan başlar.

03

Microsoft Phi-3.5 MoE Instruct Hakkında

Özet25 Haziran 2026 itibariyle

Microsoft Phi-3.5 MoE Instruct, Microsoft tarafından Metin & Sohbet kategorisinde geliştirilen bir modeldir. Microsoft Phi-3.5 MoE Instruct şu anda Railwail üzerinde kullanılamıyor. Kontekst penceresi 131.072 token içerir ve bir yanıt en fazla 4.096 token uzunluğunda olabilir.

Arka plan

Microsoft Research hakkında

Kuruluş yılı 1991 · Redmond, Washington, USA

Microsoft Research's Machine Learning Foundations group — led by Sébastien Bubeck and Ronen Eldan — drove the Phi series of small-but-capable language models. The Phi thesis is that synthetic 'textbook-quality' training data can produce small models that punch far above their weight on reasoning benchmarks. The series began with Phi-1 (1.3B, code, 2023), Phi-1.5 (general reasoning, 2023), Phi-2 (2.7B, 2023), Phi-3 (Mini, Small, Medium dense models, April 2024) and Phi-3.5 (Mini, Vision, MoE, August 2024). Phi-3.5 MoE was Microsoft's first Mixture-of-Experts Phi variant — 16 experts of 3.8B parameters each with top-2 routing. Microsoft Research itself was founded in 1991 and remains one of the largest industrial AI research organisations in the world; Phi is one of its flagship open-weights AI projects.

Microsoft Research ziyaret edin

Mimari

Mixture-of-Experts Decoder Transformer

Phi-3.5 MoE Instruct is a 16x3.8B Mixture-of-Experts decoder transformer — 16 experts each approximately the size of Phi-3-Mini, with top-2 routing yielding 6.6B active parameters out of 41.9B total. The architecture uses 32 layers, 4,096 hidden size, 32-head grouped-query attention with 8 KV heads, RoPE positional embeddings (theta=10000, extended for 128K context), SwiGLU activations, and a 32,064-token Llama-derived BPE tokeniser. Routing uses a sparse mixer with auxiliary loss for expert balancing. The model was pretrained on 4.9 trillion tokens of heavily curated data, with the Phi recipe emphasising synthetic 'textbook-quality' data generated from larger models — explicitly oversampling reasoning-dense content over breadth. Training used 512 H100 GPUs for 23 days. Post-training is supervised fine-tuning plus Direct Preference Optimisation (DPO) with explicit safety post-training. Released August 2024 under MIT license.

Parametreler
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Bağlam
131.072 token

Yetenekler

  • 16-expert MoE — Microsoft's first MoE Phi variant
  • Only 6.6B active parameters — cheap inference for MoE
  • Punches above weight: matches Mixtral 8x7B (12.9B active) and Llama 3.1 8B on many benchmarks
  • Strong math and reasoning for active-param size (MMLU 78.9, GSM8K 88.7)
  • 128K context window
  • Multilingual support for 22 languages
  • Open weights under permissive MIT license
  • Best for: cost-efficient reasoning, on-device inference (INT4 ~12GB), education and tutoring applications.

Eğitim ve lisans

Pretrained on 4.9 trillion tokens. The mix is heavily curated and includes filtered web data, synthetic 'textbook-quality' data generated from larger models, code, math and 22-language multilingual sources. Knowledge cutoff October 2023. Training used 512 NVIDIA H100 GPUs for 23 days. Post-training is supervised fine-tuning plus DPO with explicit safety post-training and red-team feedback.

Lisans: MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.

Güvenlik testleri: Microsoft published a model card and Phi-3 technical report with red-team and safety evaluation. Post-training incorporates safety alignment via DPO on red-team feedback.

Bilinen sınırlamalar

  • Total memory ~42B parameters needs ~80GB FP16 — heavier than 6.6B active suggests
  • MoE routing means latency spikes on imbalanced batches
  • Knowledge breadth narrower than larger dense models — Phi trades breadth for reasoning
  • Behind frontier models on coding benchmarks despite strong math
  • Synthetic-data-heavy training can produce 'textbook-like' answers that don't match real-world tone
  • No vision modality (use Phi-3.5-Vision instead)
04

Fiyatlandırma

Şu anda kullanılamıyor: bu model devre dışı bırakılmıştır. Bu modelin şu anda bir fiyatı yok, bu nedenle çalıştırılamıyor.

05

API

Microsoft Phi-3.5 MoE Instruct öğesini Railwail API anahtarınızla çağırın. İstekte bu model kimliğini kullanın:
phi-3-5-moe-instructAPI BelgeleriAPI Anahtarı Al

Şu anda kullanılamıyor

Modelin doğrulanmış bir fiyatı yok veya devre dışı bırakılmıştır; API çağrıları reddedilir.

06

Özellikler

Model Kimliği
phi-3-5-moe-instruct
Geliştirici
Microsoft
Giriş
Metin
Çıkış
Metin
Bağlam penceresi
131.072 token
Maks. çıkış
4.096 token
Model boyutu
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Lisans
MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.
Katalog girişi güncellendi
25 Haziran 2026

Etiketler

  • microsoft
  • open-weights
  • moe
  • multilingual
  • pricing-tbd
07

Kullanım Alanları

Nerelerde Kullanılır

  • Cost-efficient reasoning at MoE-cheap inference
  • On-device / edge AI (INT4 quantisation ~12GB)
  • Multilingual structured tasks across 22 languages
  • Education and tutoring applications
  • Math and code reasoning in resource-constrained settings
  • Self-hosted small-business assistants
08

Sık Sorulan Sorular

Microsoft Phi-3.5 MoE Instruct nedir?

Microsoft Phi-3.5 MoE Instruct, Microsoft tarafından Metin & Sohbet kategorisinde oluşturulan bir modeldir. Railwail'de listelenmiştir ancak şu anda çalıştırılamaz.

Microsoft Phi-3.5 MoE Instruct Railwail'de ne kadar maliyetlidir?

Microsoft Phi-3.5 MoE Instruct şu anda Railwail'de çalıştırılamaz, bu nedenle mevcut bir fiyat yoktur. Fiyatları olan mevcut alternatifler bu sayfanın aşağısında listelenmiştir.

Microsoft Phi-3.5 MoE Instruct'nin kontekst penceresi ne kadardır?

Microsoft Phi-3.5 MoE Instruct'nin kontekst penceresi 131.072 token içerir. Bir yanıt en fazla 4.096 token uzunluğunda olabilir.

Microsoft Phi-3.5 MoE Instruct ne kadar hızlıdır?

Railwail'de Microsoft Phi-3.5 MoE Instruct için henüz çalıştırma süresi belirtmek için yeterli ölçülen çalıştırma yoktur. Bu, giriş, ayarlar ve sağlayıcıdaki yüke bağlıdır.

Microsoft Phi-3.5 MoE Instruct, Claude Fable 5.1'dan daha iyi midir?

Bu göreve bağlıdır. Microsoft Phi-3.5 MoE Instruct (Microsoft) ve Claude Fable 5.1 (Anthropic) her ikisi de Metin & Sohbet kategorisinde modellerdir. Karşılaştırma sayfası fiyatlarını ve özelliklerini yan yana gösterir.

Microsoft Phi-3.5 MoE Instruct ve Claude Fable 5.1'ı karşılaştır

Microsoft Phi-3.5 MoE Instruct'yi şu anda kullanabilir miyim?

Şu anda kullanılamıyor: bu model devre dışı bırakılmıştır. Sayfa çevrimiçi kalır; aynı kategoriden mevcut alternatifler aşağıda listelenmiştir.

Tüm Modeller Tek Bir API Aracılığıyla

Railwail'deki her model için bir API anahtarı. Kullanım, ön ödemeli kredilerden tahsil edilir, 1 kredi = $0,01.