Microsoft Phi-3.5 MoE Instruct

Tekst & chatNiet beschikbaar
van MicrosoftModel-ID: phi-3-5-moe-instruct

Mixture-of-experts Phi-3.5: 42B total / 6.6B active params. 128k context, multilingual.

Status
Niet beschikbaar
Context
131.072 tokens
Max. uitvoer
4.096 tokens
Invoer → Uitvoer
Tekst → Tekst
Ontwikkelaar
Microsoft
Bijgewerkt
25 juni 2026

Microsoft Phi-3.5 MoE Instruct is momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

Je kunt de details op deze pagina nog steeds lezen. Kies een van de beschikbare alternatieven hieronder om direct een vergelijkbaar model uit te voeren.

Naar alternatieven
01

Vergelijkbare modellen

Alle in deze categorie
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$ 12,00/1M in

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    US$ 6,00/1M in

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$ 4,80/1M in

02

Playground

Microsoft Phi-3.5 MoE Instruct proberen

Chat

Momenteel niet beschikbaar

Momenteel niet beschikbaar: dit model is gedeactiveerd.

De playground is uitgeschakeld. Vergelijkbare modellen vind je in dezelfde categorie: Alternatieven bekijken

Microsoft Phi-3.5 MoE Instruct proberen

Stuur een bericht. Het antwoord komt volledig binnen zodra het model klaar is (geen streaming).

Systeemprompt
Max. antwoordlengte (tokens)

Deze uitvoering

Geen prijs – momenteel niet beschikbaar.

Nieuw hier?

5 gratis credits (US$ 0,05) wanneer je je aanmeldt met Google

Bruikbaar 24 uur na aanmelding, tot 5 uitvoeringen per dag en maximaal 2 credits per uitvoering. Andere aanmeldmethoden starten zonder credits.

03

Over Microsoft Phi-3.5 MoE Instruct

SamengevatPer 25 juni 2026

Microsoft Phi-3.5 MoE Instruct is een model van Microsoft in de categorie Tekst & chat. Microsoft Phi-3.5 MoE Instruct is momenteel niet beschikbaar op Railwail. Het contextvenster bevat 131.072 tokens, en een antwoord kan tot 4.096 tokens lang zijn.

Achtergrond

Over Microsoft Research

Opgericht 1991 · Redmond, Washington, USA

Microsoft Research's Machine Learning Foundations group — led by Sébastien Bubeck and Ronen Eldan — drove the Phi series of small-but-capable language models. The Phi thesis is that synthetic 'textbook-quality' training data can produce small models that punch far above their weight on reasoning benchmarks. The series began with Phi-1 (1.3B, code, 2023), Phi-1.5 (general reasoning, 2023), Phi-2 (2.7B, 2023), Phi-3 (Mini, Small, Medium dense models, April 2024) and Phi-3.5 (Mini, Vision, MoE, August 2024). Phi-3.5 MoE was Microsoft's first Mixture-of-Experts Phi variant — 16 experts of 3.8B parameters each with top-2 routing. Microsoft Research itself was founded in 1991 and remains one of the largest industrial AI research organisations in the world; Phi is one of its flagship open-weights AI projects.

Microsoft Research bezoeken

Architectuur

Mixture-of-Experts Decoder Transformer

Phi-3.5 MoE Instruct is a 16x3.8B Mixture-of-Experts decoder transformer — 16 experts each approximately the size of Phi-3-Mini, with top-2 routing yielding 6.6B active parameters out of 41.9B total. The architecture uses 32 layers, 4,096 hidden size, 32-head grouped-query attention with 8 KV heads, RoPE positional embeddings (theta=10000, extended for 128K context), SwiGLU activations, and a 32,064-token Llama-derived BPE tokeniser. Routing uses a sparse mixer with auxiliary loss for expert balancing. The model was pretrained on 4.9 trillion tokens of heavily curated data, with the Phi recipe emphasising synthetic 'textbook-quality' data generated from larger models — explicitly oversampling reasoning-dense content over breadth. Training used 512 H100 GPUs for 23 days. Post-training is supervised fine-tuning plus Direct Preference Optimisation (DPO) with explicit safety post-training. Released August 2024 under MIT license.

Parameters
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Context
131.072 tokens

Mogelijkheden

  • 16-expert MoE — Microsoft's first MoE Phi variant
  • Only 6.6B active parameters — cheap inference for MoE
  • Punches above weight: matches Mixtral 8x7B (12.9B active) and Llama 3.1 8B on many benchmarks
  • Strong math and reasoning for active-param size (MMLU 78.9, GSM8K 88.7)
  • 128K context window
  • Multilingual support for 22 languages
  • Open weights under permissive MIT license
  • Best for: cost-efficient reasoning, on-device inference (INT4 ~12GB), education and tutoring applications.

Training & licentie

Pretrained on 4.9 trillion tokens. The mix is heavily curated and includes filtered web data, synthetic 'textbook-quality' data generated from larger models, code, math and 22-language multilingual sources. Knowledge cutoff October 2023. Training used 512 NVIDIA H100 GPUs for 23 days. Post-training is supervised fine-tuning plus DPO with explicit safety post-training and red-team feedback.

Licentie: MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.

Veiligheidstests: Microsoft published a model card and Phi-3 technical report with red-team and safety evaluation. Post-training incorporates safety alignment via DPO on red-team feedback.

Bekende beperkingen

  • Total memory ~42B parameters needs ~80GB FP16 — heavier than 6.6B active suggests
  • MoE routing means latency spikes on imbalanced batches
  • Knowledge breadth narrower than larger dense models — Phi trades breadth for reasoning
  • Behind frontier models on coding benchmarks despite strong math
  • Synthetic-data-heavy training can produce 'textbook-like' answers that don't match real-world tone
  • No vision modality (use Phi-3.5-Vision instead)
04

Prijzen

Momenteel niet beschikbaar: dit model is gedeactiveerd. Er is momenteel geen prijs voor dit model, dus het kan niet worden uitgevoerd.

05

API

Roep Microsoft Phi-3.5 MoE Instruct aan met je Railwail API-sleutel. Gebruik deze model-ID in het verzoek:

Momenteel niet beschikbaar

Het model heeft geen geverifieerde prijs of is gedeactiveerd; API-aanroepen worden geweigerd.

06

Specificaties

Model-ID
phi-3-5-moe-instruct
Ontwikkelaar
Microsoft
Categorie
Tekst & chat
Invoer
Tekst
Uitvoer
Tekst
Contextvenster
131.072 tokens
Max. uitvoer
4.096 tokens
Modelgrootte
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Licentie
MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.
Catalogusitem bijgewerkt
25 juni 2026

Tags

  • microsoft
  • open-weights
  • moe
  • multilingual
  • pricing-tbd
07

Gebruiksscenario's

Waarvoor het wordt gebruikt

  • Cost-efficient reasoning at MoE-cheap inference
  • On-device / edge AI (INT4 quantisation ~12GB)
  • Multilingual structured tasks across 22 languages
  • Education and tutoring applications
  • Math and code reasoning in resource-constrained settings
  • Self-hosted small-business assistants
08

Veelgestelde vragen

Wat is Microsoft Phi-3.5 MoE Instruct?

Microsoft Phi-3.5 MoE Instruct is een model van Microsoft in de categorie Tekst & chat. Het staat in de Railwail-catalogus, maar kan momenteel niet worden uitgevoerd.

Hoeveel kost Microsoft Phi-3.5 MoE Instruct op Railwail?

Microsoft Phi-3.5 MoE Instruct kan momenteel niet op Railwail worden uitgevoerd, dus er is geen huidige prijs. Beschikbare alternatieven met prijzen staan verderop op deze pagina.

Wat is het contextvenster van Microsoft Phi-3.5 MoE Instruct?

Het contextvenster van Microsoft Phi-3.5 MoE Instruct bevat 131.072 tokens. Een antwoord kan tot 4.096 tokens lang zijn.

Hoe snel is Microsoft Phi-3.5 MoE Instruct?

Er zijn nog niet genoeg gemeten runs van Microsoft Phi-3.5 MoE Instruct op Railwail om een uitvoeringstijd op te geven. Dit hangt af van de invoer, de instellingen en de belasting bij de provider.

Is Microsoft Phi-3.5 MoE Instruct beter dan Claude Fable 5.1?

Dat hangt van de taak af. Microsoft Phi-3.5 MoE Instruct (Microsoft) en Claude Fable 5.1 (Anthropic) zijn beide modellen in de categorie Tekst & chat. De vergelijkingspagina toont hun prijzen en specificaties naast elkaar.

Microsoft Phi-3.5 MoE Instruct en Claude Fable 5.1 vergelijken

Kan ik Microsoft Phi-3.5 MoE Instruct nu gebruiken?

Momenteel niet beschikbaar: dit model is gedeactiveerd. De pagina blijft online; beschikbare alternatieven uit dezelfde categorie staan verderop.

Alle modellen via één API

Één API-sleutel voor elk model op Railwail. Gebruik wordt afgerekend via vooraf gekochte credits, 1 credit = US$ 0,01.