---
title: "Use Mac Tools, Memory, and Benchmarks"
description: "Read Noema’s live memory indicator on Mac, navigate Tools pages, and run benchmarks that use a model’s saved settings."
version: "Noema 4.1+"
platforms: ["Mac"]
reviewed: "October 3, 2026"
canonical: "https://noemaai.com/docs/mac-tools-and-performance"
---

# Use Mac Tools, Memory, and Benchmarks

The toolbar memory readout, Tools navigation, and Benchmarking Center results explained.

Noema’s Mac interface includes a compact live memory readout, Back controls for Tools pages, and benchmarks that use the selected model’s saved settings.

## Read the memory indicator

The toolbar shows Noema’s current app memory usage alongside the Mac’s total physical RAM. Hover over the indicator for an explanation.

The second value is total installed RAM, not currently free memory. The readout measures the app as a whole, including models, context, documents, and other active work.

Model file size measures storage. It is not a direct measure of runtime RAM use.

## Navigate Tools pages

Use **Back** to return from a Tools page. On supported Mac Tools pages, you can also press **Command + [**.

## Run a benchmark

1. Save the model settings you want to measure.
2. Open **Tools → Benchmarking Center**.
3. Select an installed, supported model.
4. Choose **Run Benchmark**.
5. Review the result and measurement time.

Noema uses the selected model’s saved settings and reloads the model when needed to match them.

## Understand the results

| Metric | Meaning |
| --- | --- |
| **Generation** | Output-token throughput after generation begins. Higher is faster. |
| **Prefill** | Prompt-processing throughput. |
| **First token** | Time until the first output arrives. |
| **Total time** | Duration of the measured request. |
| **Peak memory** | Highest app memory usage observed during the run. |
| **Added memory** | Increase above the benchmark’s starting memory usage. |

Depending on the runtime, token rates may use native timings or estimates.

## Compare results fairly

The benchmark uses a short request with a 512-token output limit. Everyday chats can take longer because they include conversation history, retrieved passages, tools, workspace preparation, or longer reasoning.

Generation speed excludes prompt-processing time. For comparisons, keep the model, quantization, settings, device conditions, and competing workloads as consistent as possible.

## Related documentation
- [Model Settings](https://noemaai.com/docs/model-settings)
- [Send, Retry, and Delete Messages on Mac](https://noemaai.com/docs/mac-chat)
- [Using Noema Overfit](https://noemaai.com/docs/noema-overfit)
