Start with a quick estimate using the preset or custom values. Then drill into advanced options to customize even more to fit your situation. In either case, share your results and pull in your AMD rep to discuss best options.
Max potential 3-year net savings with AMD
Choose the period of time for this analysis.
Allocate the percent of users to different workloads. Total must equal 100%.
Usage listed is tokens per user per day.
Enter the % of users per model (must total 100%). Pricing is input/output per MTok per month.
Note that models listed are for pricing assessment only and not used for AI performance comparison in the configuration recommendations.
Cloud cost is calculated from API token consumption only. No seat, subscription, or flat monthly cloud price is used.
Pricing based on publicly available data as of July 2026. Claude Sonnet includes September 2026 pricing change. Prices subject to change.
How much AI work runs locally on AMD? (0% = all cloud, 100% = all local)
Move right to reduce cloud spend. Move left to preserve access to cloud model features for complex tasks.
🖥️ AMD hardware configuration
Each device uses its own specs for throughput, power, and runtime. Enter units to deploy — a flag appears if the combined fleet falls short of your token demand.
Changing values updates Your Results.
⚡ Energy & operations
Set how long devices stay powered on. This drives your electricity cost estimate, which may differ from active AI usage hours.
☁️ Cloud API inputs
Each column below corresponds to a Cloud AI model you assigned users to above. All values are prefilled from those quantities — edit any field to override. The hybrid mix % sets how many of that model's users are served locally instead.
🧮 Cloud pricing modifiers
Prompt caching, batch, regional and long-context logic from the math concept.
Every default is illustrative for demonstration and must be validated against the latest AMD product data and current cloud pricing.
Total cost includes Cloud API costs (for Cloud and Hybrid), hardware and power costs (for Local and Hybrid) where appropriate.
Beyond cost — why teams move local
Purchase 1 unit now — one device for the local user.
Per-user capacity is calculated from separate input processing and output generation time, with a hard ceiling of one device per user.
Note: AI model performance will vary by platform so you will need to ensure that the models you want to run locally will meet your performance needs. This calculator does not consider your performance needs.
—
All values are illustrative estimates based on publicly available data · no live pricing calls are made · validate against current AMD product specifications and cloud provider pricing before making purchasing decisions.
AMD Tokenomics Calculator — illustrative cost breakdown.
Illustrative defaults. Token pricing per publicly available data as of July 2026. Prices subject to change. Power: $0.15/kWh assumed. Validate against current AMD product specs and live API pricing before making decisions.
Please read the following before using the AMD AI Cost Calculator:
Select which version to print or save as PDF.