Agnes AI logo

Agnes 3.0 Flash

Agnes AIProprietaryPending Human Review

Agnes 3.0 Flash is Agnes AI’s hosted production model for coding agents, reliable tool orchestration, long-running task execution, and instruction following. Its API supports text and image URL input with text output through Chat Completions, Responses, and Messages-compatible interfaces. The current API guide specifies 512K context and a 65,536-token output limit. The provider’s separate Preview weight card describes a 1M production configuration, so this page follows the lower documented serving contract. Production parameters and architecture are undisclosed, and the downloadable 33B Apache Preview is explicitly a different checkpoint. September month precision avoids assuming that repository publication dates describe the hosted launch.

2026-09
Undisclosed
Proprietary

Specifications

Architecture
Undisclosed
License
Proprietary
Context Window
524,288 tokens
Max Output
65,536 tokens
Type
multimodal
Modalities
textimage

Benchmark Scores

Advanced Specifications

Model Family
Agnes
API Access
Available
Chat Interface
Not Available
Variants
agnes-3.0-flash

Capabilities & Limitations

Capabilities
agentic codingreasoningtool callingimage understandinglong context
Known Limitations
Public production weights are not availableProvider sources disagree between 512K serving and 1M configuration limitsExternal application tools perform actions
Notable Use Cases
coding agentsdocument-grounded researchmulti-step enterprise workflows
Function Calling Support
Yes
Tool Use Support
Yes

Related Models