Your fleet is a team, not a GPU cluster.
The Intel machines are not useless. They can run small quantized models for background work, serve private endpoints, index your material, run checks, and keep queues moving while the M4 Pro stays responsive. Expect slower answers, not magic GPU speed.
christophers-macbook-pro
MAIN LOCAL MODELM4 Pro, 48 GB unified memory. Already runs a 35B Qwen model through llama.cpp and has a separate 9B vision model on disk. Use the 35B for focused local work; keep one large model loaded at a time. The machine has about 98 GiB free, so no model library expansion yet.
wkgd-imac
BEST INTEL MODEL HOST32 GB RAM, Intel i5-6600, Radeon graphics, 71 GB free. A realistic candidate for 3B-7B Q4 models through llama.cpp, with background-speed expectations. Also a strong index, test, and overnight queue worker.
pmd
CPU + STORAGE WORKER16 GB RAM, Xeon E5, 418 GB free, and older Radeon cards. Try 1B-3B Q4 CPU models for tags, extraction, and summaries. Better still for archives, document preparation, containers, and queued batch work than interactive chat.
wkgd-macbook
LIGHT MODEL WORKERIntel i5, 16 GB RAM, 392 GB free, with about 7 GB RAM currently available. Aim at 1B-3B Q4 models and short tasks; reserve most memory for the system and other services.
wokegods-macbook-pro
INTEL MAC MODEL NODEIntel i5, 16 GB RAM, macOS 26, 748 GB free, Metal 3. Try llama.cpp CPU inference with 3B-7B Q4; 9B may work for a focused test but will be slower. Great as a private API endpoint, model mirror, and batch worker.
wokegod-1
TINY TASK ROUTERIntel i5, 4 GB RAM, 126 GB free. Not a chat host, but useful for a lightweight dispatcher, health checks, scheduled wakeups, and perhaps a 0.5B-1B model for very short classification.
jamess-imac-pro
PENDING SHELL ACCESSOnline with SMB storage available, but Remote Login is pending. Treat it as storage until its current processor, memory, and GPU can be verified.
chris-x-imac
INTEL MODEL CANDIDATELive check now succeeds: Intel i5 3.3 GHz, 4 cores, 32 GB RAM, macOS 12.7.6, Radeon 2 GB. Try CPU-only llama.cpp with 3B-7B Q4; verify available disk before downloading. Older macOS may limit the easiest app choices.