Microsoft Phi-3.5 MoE Instruct

Texte et chatNon disponible
par MicrosoftID du modèle: phi-3-5-moe-instruct

Mixture-of-experts Phi-3.5: 42B total / 6.6B active params. 128k context, multilingual.

Statut
Non disponible
Contexte
131 072 tokens
Max. sortie
4096 tokens
Entrée → Sortie
Texte → Texte
Développeur
Microsoft
Mis à jour
25 juin 2026

Microsoft Phi-3.5 MoE Instruct n'est actuellement pas disponible

Actuellement indisponible : ce modèle a été désactivé.

Vous pouvez toujours consulter les détails sur cette page. Choisissez l'une des alternatives disponibles ci-dessous pour exécuter immédiatement un modèle comparable.

Voir les alternatives
01

Modèles comparables

Tous dans cette catégorie
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    12,00 $US/1M entrée

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    6,00 $US/1M entrée

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    4,80 $US/1M entrée

02

Playground

Essayer Microsoft Phi-3.5 MoE Instruct

Chat

Actuellement indisponible

Actuellement indisponible : ce modèle a été désactivé.

Le terrain de jeu est désactivé. Vous pouvez trouver des modèles comparables dans la même catégorie : Parcourir les alternatives

Essayer Microsoft Phi-3.5 MoE Instruct

Envoyez un message. La réponse arrive complète une fois que le modèle a terminé (sans streaming).

Prompt système
Longueur max. de la réponse (tokens)

Cette exécution

Pas de prix – actuellement indisponible.

Nouveau par ici ?

5 crédits gratuits (0,05 $US) à l'inscription avec Google

Utilisable 24 heures après l'inscription, jusqu'à 5 exécutions par jour et au maximum 2 crédits par exécution. Les autres méthodes de connexion commencent sans crédits.

03

À propos de Microsoft Phi-3.5 MoE Instruct

RésuméAu 25 juin 2026

Microsoft Phi-3.5 MoE Instruct est un modèle de Microsoft dans la catégorie Texte et chat. Microsoft Phi-3.5 MoE Instruct n'est actuellement pas disponible sur Railwail. La fenêtre de contexte contient 131 072 tokens, et une réponse peut faire jusqu'à 4096 tokens.

Arrière-plan

À propos de Microsoft Research

Fondée 1991 · Redmond, Washington, USA

Microsoft Research's Machine Learning Foundations group — led by Sébastien Bubeck and Ronen Eldan — drove the Phi series of small-but-capable language models. The Phi thesis is that synthetic 'textbook-quality' training data can produce small models that punch far above their weight on reasoning benchmarks. The series began with Phi-1 (1.3B, code, 2023), Phi-1.5 (general reasoning, 2023), Phi-2 (2.7B, 2023), Phi-3 (Mini, Small, Medium dense models, April 2024) and Phi-3.5 (Mini, Vision, MoE, August 2024). Phi-3.5 MoE was Microsoft's first Mixture-of-Experts Phi variant — 16 experts of 3.8B parameters each with top-2 routing. Microsoft Research itself was founded in 1991 and remains one of the largest industrial AI research organisations in the world; Phi is one of its flagship open-weights AI projects.

Visiter Microsoft Research

Architecture

Mixture-of-Experts Decoder Transformer

Phi-3.5 MoE Instruct is a 16x3.8B Mixture-of-Experts decoder transformer — 16 experts each approximately the size of Phi-3-Mini, with top-2 routing yielding 6.6B active parameters out of 41.9B total. The architecture uses 32 layers, 4,096 hidden size, 32-head grouped-query attention with 8 KV heads, RoPE positional embeddings (theta=10000, extended for 128K context), SwiGLU activations, and a 32,064-token Llama-derived BPE tokeniser. Routing uses a sparse mixer with auxiliary loss for expert balancing. The model was pretrained on 4.9 trillion tokens of heavily curated data, with the Phi recipe emphasising synthetic 'textbook-quality' data generated from larger models — explicitly oversampling reasoning-dense content over breadth. Training used 512 H100 GPUs for 23 days. Post-training is supervised fine-tuning plus Direct Preference Optimisation (DPO) with explicit safety post-training. Released August 2024 under MIT license.

Paramètres
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Contexte
131 072 tokens

Capacités

  • 16-expert MoE — Microsoft's first MoE Phi variant
  • Only 6.6B active parameters — cheap inference for MoE
  • Punches above weight: matches Mixtral 8x7B (12.9B active) and Llama 3.1 8B on many benchmarks
  • Strong math and reasoning for active-param size (MMLU 78.9, GSM8K 88.7)
  • 128K context window
  • Multilingual support for 22 languages
  • Open weights under permissive MIT license
  • Best for: cost-efficient reasoning, on-device inference (INT4 ~12GB), education and tutoring applications.

Entraînement et licence

Pretrained on 4.9 trillion tokens. The mix is heavily curated and includes filtered web data, synthetic 'textbook-quality' data generated from larger models, code, math and 22-language multilingual sources. Knowledge cutoff October 2023. Training used 512 NVIDIA H100 GPUs for 23 days. Post-training is supervised fine-tuning plus DPO with explicit safety post-training and red-team feedback.

Licence: MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.

Tests de sécurité: Microsoft published a model card and Phi-3 technical report with red-team and safety evaluation. Post-training incorporates safety alignment via DPO on red-team feedback.

Limitations connues

  • Total memory ~42B parameters needs ~80GB FP16 — heavier than 6.6B active suggests
  • MoE routing means latency spikes on imbalanced batches
  • Knowledge breadth narrower than larger dense models — Phi trades breadth for reasoning
  • Behind frontier models on coding benchmarks despite strong math
  • Synthetic-data-heavy training can produce 'textbook-like' answers that don't match real-world tone
  • No vision modality (use Phi-3.5-Vision instead)
04

Tarification

Actuellement indisponible : ce modèle a été désactivé. Il n'y a pas de prix pour ce modèle pour le moment, il ne peut donc pas être exécuté.

05

API

Appelez Microsoft Phi-3.5 MoE Instruct avec votre clé API Railwail. Utilisez cet ID de modèle dans la requête :

Actuellement indisponible

Le modèle n'a pas de prix vérifié ou est désactivé ; les appels API sont refusés.

06

Spécifications

ID du modèle
phi-3-5-moe-instruct
Développeur
Microsoft
Catégorie
Texte et chat
Entrée
Texte
Sortie
Texte
Fenêtre de contexte
131 072 tokens
Sortie max.
4096 tokens
Taille du modèle
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Licence
MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.
Entrée du catalogue mise à jour
25 juin 2026

Étiquettes

  • microsoft
  • open-weights
  • moe
  • multilingual
  • pricing-tbd
07

Cas d'usage

À quoi ça sert

  • Cost-efficient reasoning at MoE-cheap inference
  • On-device / edge AI (INT4 quantisation ~12GB)
  • Multilingual structured tasks across 22 languages
  • Education and tutoring applications
  • Math and code reasoning in resource-constrained settings
  • Self-hosted small-business assistants
08

Questions fréquemment posées

Qu'est-ce que Microsoft Phi-3.5 MoE Instruct ?

Microsoft Phi-3.5 MoE Instruct est un modèle de Microsoft dans la catégorie Texte et chat. Il est listé sur Railwail mais ne peut pas être exécuté pour le moment.

Combien coûte Microsoft Phi-3.5 MoE Instruct sur Railwail ?

Microsoft Phi-3.5 MoE Instruct ne peut pas être exécuté sur Railwail pour le moment, il n'y a donc pas de prix actuel. Les alternatives disponibles avec leurs prix sont listées plus bas sur cette page.

Quelle est la fenêtre de contexte de Microsoft Phi-3.5 MoE Instruct ?

La fenêtre de contexte de Microsoft Phi-3.5 MoE Instruct contient 131 072 tokens. Une réponse peut faire jusqu'à 4096 tokens.

Quelle est la vitesse de Microsoft Phi-3.5 MoE Instruct ?

Il n'y a pas encore assez d'exécutions mesurées de Microsoft Phi-3.5 MoE Instruct sur Railwail pour indiquer un temps d'exécution. Cela dépend de l'entrée, des paramètres et de la charge chez le fournisseur.

Microsoft Phi-3.5 MoE Instruct est-il meilleur que Claude Fable 5.1 ?

Cela dépend de la tâche. Microsoft Phi-3.5 MoE Instruct (Microsoft) et Claude Fable 5.1 (Anthropic) sont tous deux des modèles de la catégorie Texte et chat. La page de comparaison affiche leurs prix et spécifications côte à côte.

Comparer Microsoft Phi-3.5 MoE Instruct et Claude Fable 5.1

Puis-je utiliser Microsoft Phi-3.5 MoE Instruct maintenant ?

Actuellement indisponible : ce modèle a été désactivé. La page reste en ligne ; les alternatives disponibles de la même catégorie sont listées plus bas.

Tous les modèles via une API

Une clé API pour tous les modèles sur Railwail. L'utilisation est facturée à partir de crédits prépayés, 1 crédit = 0,01 $US.