NinjaChat
모델API요금제블로그
시작하기→
모델API요금제블로그
  1. 홈›
  2. 모델›
  3. Step 3.7 Flash

StepFun

Chat with Step 3.7 Flash online

StepFun's efficiency model — Apache 2.0, fast enough to call in a loop and capable enough to trust.

Step 3.7 Flash 사용해 보기 요금제 보기
790,000+명의 사용자가 신뢰합니다

262.144K tokens

컨텍스트

131,072 tokens

최대 출력

빠름

속도

이럴 때 활용하세요

High-volume agents

Loops where latency and price compound quickly.

Everyday coding

Fast answers on real code without the frontier price.

Search workflows

Reading and filtering a lot of material quickly.

Structured output

Clean JSON when the answer feeds a pipeline.

Prompts to steal

Tuned to this model — click any line to copy.

Step 3.7 Flash 소개

Step 3.7 Flash is StepFun's efficiency model, released on 28 May 2026 under an Apache 2.0 licence. It is a sparse mixture-of-experts design of about 198 billion parameters that activates only around 11 billion per token, which is the whole point: it answers quickly and costs little, while staying capable enough for real work.

StepFun built it for coding agents and search-style workflows — the kind of job where a model is called many times in a loop and both latency and price compound. Throughput runs to several hundred tokens a second on a good rail, and the model exposes selectable reasoning levels so a caller can trade depth against speed per request rather than being locked into one setting. It takes tools and tool_choice for function calling and supports a JSON response format when the output feeds a pipeline.

The value claim is the interesting part. On SWE-Bench Verified with its advisor mode enabled, StepFun reports Step 3.7 Flash reaching around 97% of a frontier model's coding score at roughly a ninth of the per-task cost. Treat any single benchmark carefully, but the shape of the claim matches what the architecture is for.

The context window is 262,144 tokens, so a substantial codebase or a long document set fits in one conversation. On NinjaChat it is included in every plan — a good default when you want a capable model to answer fast and often.

Step 3.7 Flash 사용 방법

1분 안에 첫 결과물까지.

01

Create a NinjaChat account and choose a plan

02

Open chat and pick Step 3.7 Flash from the model list

03

Use it as your default for fast, frequent questions

04

Escalate to a frontier model when a task gets hard

더 나은 결과를 위한 팁

  • Great as the first pass before a bigger model
  • Ask for JSON when the answer feeds something else
  • Its context is large; paste the whole file
  • Pair it with a frontier model rather than replacing one

Step 3.7 Flash vs 다른 대안들

{model}이 앞서는 부분과 그렇지 않은 부분, 솔직하게 비교합니다.

vs Qwen 3.8 Flash

모델 보기

+Larger context and selectable reasoning levels

–Qwen 3.8 Flash is cheaper still

vs GLM 4.7

모델 보기

+Cheaper per token at a similar everyday coding job

–GLM 4.7 has the wider GLM ecosystem behind it

vs Kimi K2.7 Code

모델 보기

+Much faster and cheaper for routine work

–Kimi K2.7 Code is far stronger on long agent runs

자주 묻는 질문

채팅 그 이상 — 이미지 & 동영상

모든 NinjaChat 요금제에는 50개 이상의 모델, 완전한 이미지 스튜디오, 동영상 생성이 포함됩니다.

A real FLUX Pro Ultra outputA real Google Imagen 4 outputA real Seedream output
전체 모델 보기 →

Step 3.7 Flash, 그리고 50개의 모델 더.

구독 하나로 NinjaChat의 모든 모델을 이용할 수 있습니다. Step 3.7 Flash도 포함됩니다.

시작하기요금제 비교
Download on the App Store

더 많은 모델 둘러보기

Qwen 3.8 Flash

알리바바의 최신 멀티모달 추론 모델, 빠르고 저렴합니다

→

GLM 4.7

코딩, 추론, 범용 작업에 강한 Zhipu의 효율적인 모델

→

Kimi K2.7 Code

Moonshot's coding specialist for long agent runs

→
NinjaChat

모든 AI. 하나의 앱.

App Store에서 다운로드

제품

  • 대시보드
  • ninja
  • 요금제
  • Enterprise
  • 무료 AI 도구
  • Cinema
  • 제휴 프로그램
  • iOS 앱

개발자

  • API
  • API 모델
  • Router
  • MCP / Agents
  • Search API
  • API Pricing
  • Migration Guides
  • API 문서

모델

  • Model Council
  • Seed 1.8
  • Gemini 2.5 Flash
  • Gemini 2.5 Pro
  • Gemini 3 Flash Preview
  • 전체 모델 보기

회사

  • 블로그
  • 커뮤니티
  • 채용
  • 지원
  • 개인정보처리방침
  • 이용약관
  • Safety Protocol
  • Do Not Sell or Share My Personal Information

Copyright © 2026 NinjaChat AI. 제공 Bloon 모든 권리 보유.