Agnes 3.0 Flash
Agnes 3.0 Flash is Agnes AI’s hosted production model for coding agents, reliable tool orchestration, long-running task execution, and instruction following. Its API supports text and image URL input with text output through Chat Completions, Responses, and Messages-compatible interfaces. The current API guide specifies 512K context and a 65,536-token output limit. The provider’s separate Preview weight card describes a 1M production configuration, so this page follows the lower documented serving contract. Production parameters and architecture are undisclosed, and the downloadable 33B Apache Preview is explicitly a different checkpoint. September month precision avoids assuming that repository publication dates describe the hosted launch.
Specifications
- Architecture
- Undisclosed
- License
- Proprietary
- Context Window
- 524,288 tokens
- Max Output
- 65,536 tokens
- Type
- multimodal
- Modalities
- textimage
Benchmark Scores
Advanced Specifications
- Model Family
- Agnes
- API Access
- Available
- Chat Interface
- Not Available
- Variants
- agnes-3.0-flash
Capabilities & Limitations
- Capabilities
- agentic codingreasoningtool callingimage understandinglong context
- Known Limitations
- Public production weights are not availableProvider sources disagree between 512K serving and 1M configuration limitsExternal application tools perform actions
- Notable Use Cases
- coding agentsdocument-grounded researchmulti-step enterprise workflows
- Function Calling Support
- Yes
- Tool Use Support
- Yes