Power rating
Published detail480 W maximum continuous power rating. This is not typical consumption. [1]
Manufacturer product image. The selected configuration may vary.
Apple M5 Ultra · 30-core CPU / 64-core GPU
An M5 Ultra preorder configuration for exploring local inference in the Metal and MLX ecosystem. Exact runtime support remains to be checked.
Compared withNVIDIA DGX SparkFramework Desktop · Ryzen AI Max+Apple Mac Studio · M3 Ultra
BEFORE IT ARRIVES
What this configuration needs on your premises.
An office desktop candidate with published electrical limits. Apple’s wall-power test uses a larger configuration than this entry; measure your workload before budgeting a fleet.
480 W maximum continuous power rating. This is not typical consumption. [1]
100–240 V AC, 50–60 Hz, single phase; supplied AC power cord. [1]
No machine-specific branch-circuit requirement verified. Check the nameplate, regional cord and other loads on the same circuit.
Active fan cooling. Apple requires a hard, stable surface with air circulation beneath and around the computer; keep ventilation openings clear. [3][4]
10–35°C operating ambient, 5–90% relative humidity, noncondensing. [1]
No comparable noise measurement under an AI workload verified. Ask for dB(A), distance, workload and ambient conditions; audition it before placing it beside people.
ACTUAL CONSUMPTION
PSU capacity, chip TDP and measured consumption describe different things. Only measurements with a stated configuration and test method appear here.
M5 Ultra · 36-core CPU / 80-core GPU · 512 GB · 16 TB SSD
Apple measured at the wall, including power-supply losses, at 20.2°C. Idle: Finder only. Maximum: compute-intensive processor workload. External peripherals excluded. This is a different configuration from this catalogue entry, not an LLM benchmark or a guarantee for your workload. [2]
Plan for the computer and its power supply to heat the room. Check ventilation during long jobs and when several machines share a small office.
Lab planning guidance. Office suitability depends on the installed configuration, shared circuits, heat removal and acceptable noise; a workstation label alone does not settle it.
Enter average power at the wall for the hours you are modeling, or an explicit planning assumption. Supply wattage and chip TDP are not consumption measurements.
Add wall power to calculate a scenario.
Scenario only. Hours outside this period are excluded: include idle time for an always-on system. Displays, networking, UPS losses and room air conditioning are extra. Heat assumes the computer and power supply release their heat into this room; it does not size a circuit or an HVAC system. Currency selects your tariff’s unit, with no conversion.
Manufacturer documentation and attributed first-hand hardware inspections, scoped to the named model or configuration. “Not verified” identifies a gap in this review, not a claim that the manufacturer has no documentation. Installation assessments are the Lab’s interpretation.
Keep sensitive processing on your infrastructure. Inspect every connection in the complete stack.
Control access, connectivity and updates. A local machine still needs a secure deployment.
Generate tokens on your own compute. Understand the full cost of the useful work it delivers.
SHARED WORK AROUND THIS PLATFORM
Reference-platform guidance and field records. Exact OEM configurations need their own validation.
THE PLATFORM BEHIND THE MACHINE
2 catalog configurations
BEYOND THE SPECIFICATION
Published guidance and checklists to help you investigate the full deployment.
An execution agent and confidential engineering test assets show how on-premises AI extends beyond office applications.
Local model inference is one part of the path. Test retrieval, documents, identity, tools and restart behavior as well.
Put quality, retries, utilization and operating effort beside token throughput.
THE CONFIGURATION
ON YOUR PREMISES
BEFORE YOU CHOOSE
Apple store lists this configuration for pre-order, available from 22 September. No usable price was exposed in the reviewed store page.
This is a new configuration. Runtime compatibility, actual GPU memory allowance and performance have not been validated by this lab.
Manufacturer specificationsHelp improve this record. Suggest a correction
IN THE FIELD
Operating records from customers, contributors and the Lab.
In an office, a factory, an engineering lab or beside your test equipment. Share a deployment and keep the people behind it credited.
Contribute a deploymentTHE SHARED TECHNICAL LIBRARY
Supplier recipes, community experience and UT’s deployment lessons, connected to the hardware.
Text generation and fine-tuning on Apple silicon.
Use a compatible model conversion. Measure memory and latency at the context length your team needs.
Examine what still connects beyond the premises.
Distinguish initial installation, inference, application use, authentication and updates.
Confidential Wi-Fi module test automation.
Exact hardware and proprietary test assets are not disclosed. This is an application field record.
Loading updates…
MEMBER RATINGS · SELF-DECLARED, NOT MEASURED
One rating per member, tied to a Lab profile. Stars count immediately; written verdicts are published after a quick review by the Lab team. A rating is an opinion about fit for a job, not a benchmark.
THE PEOPLE BEHIND THE MACHINES
Questions, ideas and experience from putting AI machines to work. You do not need a finished deployment to contribute.
Loading the discussion…
THE MACHINE × THE MODEL × THE WORKLOAD
Connect the model, serving stack, workload and cost of operating this box.
| Model | Serving configuration | Capacity planning | Evidence | Inspect |
|---|---|---|---|---|
| 4-bit · MLX | Maintainer example | |||
| 4-bit · MLX | Maintainer example | |||
| 4-bit · MLX | Community model card |
Ecosystem-level guidance; no M3 Ultra or M5 Ultra reproduction is attached.
Available GPU memory and runtime support must be checked on the exact Mac.
Each measured result belongs to one box configuration, one model, one serving stack and one workload. Publish latency, per-stream speed, request rate, errors and task quality together.