Cerebras-Llama-4-Scout-17B-16E-Instruct

Cerebras-Llama-4-Scout-17B-16E-Instruct: a llama model from llama, ~128K context, knowledge cutoff 2025-01

Published
scheduleReleasedApril 5, 2025

Cerebras-Llama-4-Scout-17B-16E-Instruct is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

starsCapabilities

codeFunction calling

paymentsContext and pricing

Context limit128,000
Max output4,096
Knowledge cutoff2025-01
Input price$0/ 1M tokens
Output price$0/ 1M tokens

descriptionOverview

Overview

Cerebras-Llama-4-Scout-17B-16E-Instruct is provided by llama, model ID cerebras-llama-4-scout-17b-16e-instruct

Key specs

  • Context: 128K tokens
  • Max output: 4.1K tokens
  • Knowledge cutoff: 2025-01
  • Input price: $0/1M
  • Output price: $0/1M

Best for

Consider Cerebras-Llama-4-Scout-17B-16E-Instruct when comparing context length, pricing, multimodal support and relay availability

lightbulbUse cases

  • Assistants and customer support
  • Content generation and rewriting
  • Knowledge Q&A and summarization
  • Structured extraction

thumb_upStrengths

  • Large context window (~128K)
  • Open weights, can be self-hosted

infoLimitations

  • Knowledge cutoff 2025-01; newer facts need external retrieval
  • Pricing and availability vary by upstream and relay; verify with official docs and tests

linkReferences

This content is compiled from official documentation and public sources. Always refer to official documentation for final details