Step 3.5 Flash is Stepfun's lightweight model designed for fast text generation, featuring a 262K token context window and high throughput performance.
| Provider | Input / 1M | Output / 1M |
|---|---|---|
| $0.100 | $0.300 |
Prices updated daily. Last check: Sep 6, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
Benchmarks measured Sep 2026. Scores are independent evaluations, not vendor-reported.
Step 3.5 Flash is designed for applications requiring fast text generation and high throughput processing. Its large context window combined with quick generation speeds makes it well-suited for document summarization, content processing pipelines, customer service automation, and real-time text analysis. The model works effectively for workflows that need to process substantial text volumes quickly, such as content moderation, text classification at scale, or generating responses in chat applications where speed is prioritized. Its lightweight nature makes it cost-effective for high-volume deployments where complex reasoning capabilities are not required.
Step 3.5 Flash pricing varies by provider and may include different rates for input and output tokens. Check the pricing table above for current rates across all available providers offering this model.
Step 3.5 Flash excels at high-throughput text processing tasks where speed is important. With its 169.5 tokens per second generation rate and large 262K context window, it's ideal for document summarization, content processing pipelines, customer service automation, and real-time text analysis where quick responses matter more than complex reasoning.
No, Step 3.5 Flash is designed as a streamlined text-only model without tool calling capabilities or support for images and other modalities. This focused approach contributes to its fast performance characteristics and makes it suitable for pure text generation tasks.