Microsoft Phi-3.5 MoE Instruct

Texto e chatIndisponível
por MicrosoftID do modelo: phi-3-5-moe-instruct

Mixture-of-experts Phi-3.5: 42B total / 6.6B active params. 128k context, multilingual.

Status
Indisponível
Contexto
131.072 tokens
Saída máxima
4.096 tokens
Entrada → Saída
Texto → Texto
Desenvolvedor
Microsoft
Atualizado
25 de junho de 2026

Microsoft Phi-3.5 MoE Instruct não está disponível no momento

Atualmente indisponível: este modelo foi desativado.

Você ainda pode ler os detalhes nesta página. Escolha uma das alternativas disponíveis abaixo para executar um modelo comparável imediatamente.

Ir para alternativas
01

Modelos comparáveis

Todos nesta categoria
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$ 12,00/1M entrada

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    US$ 6,00/1M entrada

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$ 4,80/1M entrada

02

Playground

Experimentar Microsoft Phi-3.5 MoE Instruct

Chat

Indisponível no momento

Atualmente indisponível: este modelo foi desativado.

O playground está desativado. Encontre modelos comparáveis na mesma categoria: Ver alternativas

Experimentar Microsoft Phi-3.5 MoE Instruct

Envie uma mensagem. A resposta chega completa assim que o modelo termina (sem streaming).

Prompt do sistema
Comprimento máximo da resposta (tokens)

Esta execução

Sem preço – indisponível no momento.

Novo por aqui?

5 créditos grátis (US$ 0,05) ao se inscrever com Google

Utilizável 24 horas após inscrição, até 5 execuções por dia e no máximo 2 créditos por execução. Outros métodos de login começam sem créditos.

03

Sobre Microsoft Phi-3.5 MoE Instruct

ResumoA partir de 25 de junho de 2026

Microsoft Phi-3.5 MoE Instruct é um modelo de Microsoft na categoria Texto e chat. Microsoft Phi-3.5 MoE Instruct não está disponível no Railwail no momento. A janela de contexto contém 131.072 tokens, e uma resposta pode ter até 4.096 tokens.

Fundo

Sobre Microsoft Research

Fundado em 1991 · Redmond, Washington, USA

Microsoft Research's Machine Learning Foundations group — led by Sébastien Bubeck and Ronen Eldan — drove the Phi series of small-but-capable language models. The Phi thesis is that synthetic 'textbook-quality' training data can produce small models that punch far above their weight on reasoning benchmarks. The series began with Phi-1 (1.3B, code, 2023), Phi-1.5 (general reasoning, 2023), Phi-2 (2.7B, 2023), Phi-3 (Mini, Small, Medium dense models, April 2024) and Phi-3.5 (Mini, Vision, MoE, August 2024). Phi-3.5 MoE was Microsoft's first Mixture-of-Experts Phi variant — 16 experts of 3.8B parameters each with top-2 routing. Microsoft Research itself was founded in 1991 and remains one of the largest industrial AI research organisations in the world; Phi is one of its flagship open-weights AI projects.

Visite Microsoft Research

Arquitetura

Mixture-of-Experts Decoder Transformer

Phi-3.5 MoE Instruct is a 16x3.8B Mixture-of-Experts decoder transformer — 16 experts each approximately the size of Phi-3-Mini, with top-2 routing yielding 6.6B active parameters out of 41.9B total. The architecture uses 32 layers, 4,096 hidden size, 32-head grouped-query attention with 8 KV heads, RoPE positional embeddings (theta=10000, extended for 128K context), SwiGLU activations, and a 32,064-token Llama-derived BPE tokeniser. Routing uses a sparse mixer with auxiliary loss for expert balancing. The model was pretrained on 4.9 trillion tokens of heavily curated data, with the Phi recipe emphasising synthetic 'textbook-quality' data generated from larger models — explicitly oversampling reasoning-dense content over breadth. Training used 512 H100 GPUs for 23 days. Post-training is supervised fine-tuning plus Direct Preference Optimisation (DPO) with explicit safety post-training. Released August 2024 under MIT license.

Parâmetros
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Contexto
131.072 tokens

Capacidades

  • 16-expert MoE — Microsoft's first MoE Phi variant
  • Only 6.6B active parameters — cheap inference for MoE
  • Punches above weight: matches Mixtral 8x7B (12.9B active) and Llama 3.1 8B on many benchmarks
  • Strong math and reasoning for active-param size (MMLU 78.9, GSM8K 88.7)
  • 128K context window
  • Multilingual support for 22 languages
  • Open weights under permissive MIT license
  • Best for: cost-efficient reasoning, on-device inference (INT4 ~12GB), education and tutoring applications.

Treinamento & licença

Pretrained on 4.9 trillion tokens. The mix is heavily curated and includes filtered web data, synthetic 'textbook-quality' data generated from larger models, code, math and 22-language multilingual sources. Knowledge cutoff October 2023. Training used 512 NVIDIA H100 GPUs for 23 days. Post-training is supervised fine-tuning plus DPO with explicit safety post-training and red-team feedback.

Licença: MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.

Testes de segurança: Microsoft published a model card and Phi-3 technical report with red-team and safety evaluation. Post-training incorporates safety alignment via DPO on red-team feedback.

Limitações conhecidas

  • Total memory ~42B parameters needs ~80GB FP16 — heavier than 6.6B active suggests
  • MoE routing means latency spikes on imbalanced batches
  • Knowledge breadth narrower than larger dense models — Phi trades breadth for reasoning
  • Behind frontier models on coding benchmarks despite strong math
  • Synthetic-data-heavy training can produce 'textbook-like' answers that don't match real-world tone
  • No vision modality (use Phi-3.5-Vision instead)
04

Preços

Atualmente indisponível: este modelo foi desativado. Não há preço para este modelo no momento, portanto não pode ser executado.

05

API

Chame Microsoft Phi-3.5 MoE Instruct com sua chave de API Railwail. Use este ID de modelo na solicitação:

Indisponível no momento

O modelo não tem preço verificado ou está desativado; chamadas de API são recusadas.

06

Especificações

ID do modelo
phi-3-5-moe-instruct
Desenvolvedor
Microsoft
Categoria
Texto e chat
Entrada
Texto
Saída
Texto
Janela de contexto
131.072 tokens
Saída máx.
4.096 tokens
Tamanho do modelo
41.9B total, 6.6B active per token (16 experts of ~3.8B each, top-2 routing)
Licença
MIT License for the open weights. Commercial use, redistribution and modification permitted without restriction — one of the most permissive licenses among major open-weight LLMs.
Entrada do catálogo atualizada
25 de junho de 2026

Etiquetas

  • microsoft
  • open-weights
  • moe
  • multilingual
  • pricing-tbd
07

Casos de uso

Para que é utilizado

  • Cost-efficient reasoning at MoE-cheap inference
  • On-device / edge AI (INT4 quantisation ~12GB)
  • Multilingual structured tasks across 22 languages
  • Education and tutoring applications
  • Math and code reasoning in resource-constrained settings
  • Self-hosted small-business assistants
08

Perguntas frequentes

O que é Microsoft Phi-3.5 MoE Instruct?

Microsoft Phi-3.5 MoE Instruct é um modelo de Microsoft na categoria Texto e chat. Está listado no Railwail, mas não pode ser executado no momento.

Quanto custa Microsoft Phi-3.5 MoE Instruct no Railwail?

Microsoft Phi-3.5 MoE Instruct não pode ser executado no Railwail no momento, portanto não há preço atual. Alternativas disponíveis com preços estão listadas mais abaixo nesta página.

Qual é a janela de contexto de Microsoft Phi-3.5 MoE Instruct?

A janela de contexto de Microsoft Phi-3.5 MoE Instruct contém 131.072 tokens. Uma resposta pode ter até 4.096 tokens.

Qual é a velocidade de Microsoft Phi-3.5 MoE Instruct?

Ainda não há execuções medidas suficientes de Microsoft Phi-3.5 MoE Instruct no Railwail para indicar um tempo de execução. Depende da entrada, das configurações e da carga no provedor.

Microsoft Phi-3.5 MoE Instruct é melhor que Claude Fable 5.1?

Depende da tarefa. Microsoft Phi-3.5 MoE Instruct (Microsoft) e Claude Fable 5.1 (Anthropic) são ambos modelos na categoria Texto e chat. A página de comparação mostra seus preços e especificações lado a lado.

Comparar Microsoft Phi-3.5 MoE Instruct e Claude Fable 5.1

Posso usar Microsoft Phi-3.5 MoE Instruct agora?

Atualmente indisponível: este modelo foi desativado. A página permanece online; alternativas disponíveis da mesma categoria estão listadas mais abaixo.

Todos os modelos através de uma API

Uma chave API para todos os modelos no Railwail. O uso é cobrado a partir de créditos pré-pagos, 1 crédito = US$ 0,01.