Agnes 3.0 Flash is Agnes AI’s hosted production model for coding agents, reliable tool orchestration, long-running task execution, and instruction following. Its API supports text and image URL input with text output through Chat Completions, Responses, and Messages-compatible interfaces. The current API guide specifies 512K context and a 65,536-token output limit. The provider’s separate Preview weight card describes a 1M production configuration, so this page follows the lower documented serving contract. Production parameters and architecture are undisclosed, and the downloadable 33B Apache Preview is explicitly a different checkpoint. September month precision avoids assuming that repository publication dates describe the hosted launch.
Typemultimodal
ParametersUnknown