Overview
A high-efficiency vision-language Flash model for real-world agents, with about 198B total parameters, about 11B active, 256K context, image understanding, and tool orchestration.
A high-efficiency vision-language Flash model for real-world agents, with about 198B total parameters, about 11B active, 256K context, image understanding, and tool orchestration.
Only published specifications are shown.
A high-efficiency vision-language Flash model for real-world agents, with about 198B total parameters, about 11B active, 256K context, image understanding, and tool orchestration.
Step 3.7 Flash extends Step 3.5 Flash’s high-throughput agent path to native vision, search, and longer-horizon tool execution.
Step 3.7 Flash launched with open weights and deployment materials.
View sourceSelect a card to view the related model.
Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.