NinjaChat
模型API定价博客
开始使用→
模型API定价博客
  1. 首页›
  2. 模型›
  3. Step 3.7 Flash

StepFun

Chat with Step 3.7 Flash online

StepFun's efficiency model — Apache 2.0, fast enough to call in a loop and capable enough to trust.

试用 Step 3.7 Flash 查看方案
790,000+ 位用户的信赖

262.144K tokens

上下文窗口

131,072 tokens

最大输出

快速

速度

你会用它做什么

High-volume agents

Loops where latency and price compound quickly.

Everyday coding

Fast answers on real code without the frontier price.

Search workflows

Reading and filtering a lot of material quickly.

Structured output

Clean JSON when the answer feeds a pipeline.

Prompts to steal

Tuned to this model — click any line to copy.

关于 Step 3.7 Flash

Step 3.7 Flash is StepFun's efficiency model, released on 28 May 2026 under an Apache 2.0 licence. It is a sparse mixture-of-experts design of about 198 billion parameters that activates only around 11 billion per token, which is the whole point: it answers quickly and costs little, while staying capable enough for real work.

StepFun built it for coding agents and search-style workflows — the kind of job where a model is called many times in a loop and both latency and price compound. Throughput runs to several hundred tokens a second on a good rail, and the model exposes selectable reasoning levels so a caller can trade depth against speed per request rather than being locked into one setting. It takes tools and tool_choice for function calling and supports a JSON response format when the output feeds a pipeline.

The value claim is the interesting part. On SWE-Bench Verified with its advisor mode enabled, StepFun reports Step 3.7 Flash reaching around 97% of a frontier model's coding score at roughly a ninth of the per-task cost. Treat any single benchmark carefully, but the shape of the claim matches what the architecture is for.

The context window is 262,144 tokens, so a substantial codebase or a long document set fits in one conversation. On NinjaChat it is included in every plan — a good default when you want a capable model to answer fast and often.

如何使用 Step 3.7 Flash

从零到第一个结果,不到一分钟。

01

Create a NinjaChat account and choose a plan

02

Open chat and pick Step 3.7 Flash from the model list

03

Use it as your default for fast, frequent questions

04

Escalate to a frontier model when a task gets hard

获得更好结果的技巧

  • Great as the first pass before a bigger model
  • Ask for JSON when the answer feeds something else
  • Its context is large; paste the whole file
  • Pair it with a frontier model rather than replacing one

Step 3.7 Flash 与同类模型对比

客观对比——{model} 的优势所在,以及它的不足之处。

对比 Qwen 3.8 Flash

查看模型

+Larger context and selectable reasoning levels

–Qwen 3.8 Flash is cheaper still

对比 GLM 4.7

查看模型

+Cheaper per token at a similar everyday coding job

–GLM 4.7 has the wider GLM ecosystem behind it

对比 Kimi K2.7 Code

查看模型

+Much faster and cheaper for routine work

–Kimi K2.7 Code is far stronger on long agent runs

常见问题解答

不止于对话——图像与视频

每个 NinjaChat 方案均包含 50+ 个模型、完整图像工作室和视频生成功能。

A real FLUX Pro Ultra outputA real Google Imagen 4 outputA real Seedream output
查看所有模型 →

Step 3.7 Flash,还有另外 50 多个模型。

一个订阅即可使用 NinjaChat 上的所有模型,包括 Step 3.7 Flash。

开始使用对比方案
Download on the App Store

更多模型,等你探索

Qwen 3.8 Flash

阿里巴巴最新的多模态推理模型,快速且低价

→

GLM 4.7

智谱推出的高效模型,擅长编程、推理与通用任务

→

Kimi K2.7 Code

Moonshot's coding specialist for long agent runs

→
NinjaChat

所有 AI。 一个应用。

在 App Store 上下载

产品

  • 控制台
  • ninja
  • 定价
  • Enterprise
  • 免费 AI 工具
  • Cinema
  • 推广合作
  • iOS 应用

开发者

  • API
  • API 模型
  • Router
  • MCP / Agents
  • Search API
  • API Pricing
  • Migration Guides
  • API 文档

模型

  • Model Council
  • Seed 1.8
  • Gemini 2.5 Flash
  • Gemini 2.5 Pro
  • Gemini 3 Flash Preview
  • 查看全部模型

公司

  • 博客
  • 社区
  • 加入我们
  • 客服支持
  • 隐私政策
  • 服务条款
  • Safety Protocol
  • Do Not Sell or Share My Personal Information

版权所有 © 2026 NinjaChat AI。 旗下产品 Bloon 保留所有权利。