Síntesis de evidencia pública

Adaptación del modelo local para Mac mini y Mac Studio

Una herramienta de planificación para los últimos representantes de peso abierto Qwen3.8, DeepSeek V4, GLM-5.3 y Gemma 4. Comience con datos de archivos públicos y cálculos de capacidad, tiempo de ejecución separado y evidencia de trabajo, y luego decida si se justifica un reemplazo.

Última revisión: 28 de agosto de 2026 · Orientación independiente, no consejo de Apple

Conclusión breve

Esta página compara los representantes de cada familia que actualmente son de interés. Las versiones anteriores se conservan únicamente como registro de compatibilidad y se excluyen de la matriz principal. Incluso si el modelo cabe en la memoria, es posible que se desconozcan la invocación de herramientas, JSON, el comportamiento visual, la concurrencia y la persistencia.

herramientas de dialogo

Comprueba si el modelo es compatible con tu Mac.

Elija su modelo o configuración de Mac para ver los presupuestos ponderados, el margen de memoria, la evidencia del tiempo de ejecución y los próximos pasos para nuestro Asesor gratuito.

Explorador de cumplimiento abierto
Síntesis de evidencia pública, no un punto de referencia propio.

Esta herramienta de planificación combina tarjetas modelo oficiales, metadatos de archivos públicos GGUF/QAT, documentación en tiempo de ejecución, evaluaciones publicadas e informes comunitarios claramente etiquetados. No afirma que Keep o Upgrade ejecutaron todos los modelos, y Apple o los proveedores de modelos no respaldan esta página.

Descargar CSVDescargar JSON · Los mismos datos de origen que esta página.

Las descargas son instantáneas fáciles de citar con URL de origen, niveles de evidencia y fecha de los datos. Enlace a este Hub como explicación canónica; los puntos finales de descarga permanecen fuera del mapa del sitio.

Mac mini M6 · 16GB153GB/s · US$899 starting price

Matriz de evidencia modelo

Modelos de peso abierto más recientes y de mayor interés

La matriz principal prioriza los lanzamientos emblemáticos actuales y los representantes útiles de planificación de Mac. La capacidad utiliza el tamaño del artefacto público seleccionado cuando esté disponible; Las cifras basadas en parámetros son únicamente estimaciones de planificación. Expanda Fuentes y límites en cualquier fila para inspeccionar señales de red atribuibles y sus límites.

Familia / nivelModelo y versiónCuantización / artefactoCapacidad en Mac seleccionadoEvidencia en tiempo de ejecuciónConfidenceDetails
QwenPequeñoQwen3.8 27BQwen3.8 · 27B denseQ4_0 · GGUF16.1 GB · Public file metadataNo cabe-8.06GB restantes tras la reservaSoporte oficialmlx2 señales sociales · ver fuenteshighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public file metadata16.1 GB · exact public file metadata
  • Fuente del runtimemlx · Soporte oficial
  • Evidencia de la comunidad (2)Los informes públicos son solo señales; nunca convierten esta página en una prueba del proyecto.
  • Reddit · Qwen3.8-27B on a 24GB M4 Pro Mac miniMac mini · M4 Pro · 24GB · llama.cpp b10488 · GGUF Q4_K_M / IQ4_XSQ4_K_M 17.77GB; about 11.4 tok/s decode; about 16.6GB resident with q8 KV at 32K context · medium confidenceA first-person report says the 27B model is usable on 24GB, but the memory ceiling is visible once context and KV cache are included. Limitación: One user, one build, and one workload; not a Keep or Upgrade test.
  • Tech media / blog · Qwen3.8-27B 4-bit on a 32GB M1 ProMacBook Pro · M1 Pro · 32GB · MLX-VLM · MLX 4-bit8–8.7 tok/s; weights about 16.05GB; peak memory about 18.5–21.7GB · medium confidenceA detailed local deployment log reports a stable 32GB starting point and explicitly discourages 16GB for this 27B representative. Limitación: Community log with different thermals, context, and prompt mix from the product matrix.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
QwenPrincipalQwen3.8-Flash-Next 180B (6B active)Qwen3.8-Flash-Next · 180B total / 6B activeBF16 · Safetensors360.0 GB · Public file metadataNo cabe-352GB restantes tras la reservaInforme comunitariomlx1 señal social · ver fuentesmediumComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public file metadata360.0 GB · exact public file metadata
  • Fuente del runtimemlx · Informe comunitario
  • Evidencia de la comunidad (1)Los informes públicos son solo señales; nunca convierten esta página en una prueba del proyecto.
  • Reddit · Latest-model comparison thread (non-Apple hardware)4× DGX Spark / other non-Apple systemsUsers discuss Qwen3.8 Flash as an exploration/subagent model; no Mac tuple or reproducible throughput is provided. · low confidenceThis is a freshness and interest signal only, not evidence that Flash Next fits a Mac configuration. Limitación: Cross-hardware discussion; excluded from Mac capacity conclusions.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
QwenLímiteQwen3.8 2.4T-A95BQwen3.8 · 2400B total / 95B activeQ4_K_M · GGUF1440.0 GB · Public model source; exact file not registeredNo cabe-1432GB restantes tras la reservaSoporte oficialmlxmediumComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public model source; exact file not registered1440.0 GB · parameter planning estimate
  • Fuente del runtimemlx · Soporte oficial
  • No se ha capturado un informe fiable de primera mano sobre Mac en esta actualización.La capacidad sigue siendo una estimación de planificación basada en metadatos públicos; la ausencia de pruebas no valida el runtime.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
DeepSeekPrincipalDeepSeek-V4-Flash 284B-A13BDeepSeek V4 · 284B total / 13B activeMLX-4bit · MLX170.4 GB · Public model source; exact file not registeredNo cabe-162.4GB restantes tras la reservaInforme comunitariomlx7 señales sociales · ver fuenteshighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public model source; exact file not registered170.4 GB · parameter planning estimate
  • Fuente del runtimemlx · Informe comunitario
  • Evidencia de la comunidad (7)Los informes públicos son solo señales; nunca convierten esta página en una prueba del proyecto.
  • Reddit · DeepSeek V4 Flash on an M5 Max 128GBMac · Apple M5 Max · 128GB · ds4 / SSD-streaming path · DwarfStar IQ2XXSAbout 81GB model working set and about 31.06 tok/s in the reported run · medium confidenceA direct Apple Silicon report shows a practical 128GB path for the Flash representative. Limitación: Community benchmark; exact context, thermal state, and build can change the result.
  • Tech media / blog · DeepSeek V4 Flash on an M3 Max 128GBMacBook Pro · M3 Max · 128GB · ds4 experimental forkAbout 21 tok/s and about 81GB resident in the reported run · medium confidenceAn independent technical write-up corroborates that 128GB-class Apple Silicon can run a local Flash path. Limitación: The article is tied to an experimental fork and does not validate mainline Ollama or llama.cpp.
  • Reddit · DeepSeek V4 Flash on 64GB Macs with SSD streaming64GB Apple Silicon Macs · 64GB · ds4 SSD streaming · IQ2XXS / 2-bit classReports cluster around roughly 10–15 tok/s; technically possible, but not a comfortable in-memory workflow · medium confidence64GB appears technically viable only with aggressive quantization and storage streaming, so it is kept as a constrained edge case. Limitación: Mixed community reports and different Mac generations; do not read this as a 64GB recommendation.
  • GitHub · MLX DeepSeek V4 residency growth and resource-limit crashM4 Max 128GB and other Apple Silicon reports · 128GB · MLX-LMReproduction reports deterministic resource-limit failure around 11,300 generated tokens · high confidenceThe issue is important negative evidence: a model can fit and still fail on long generations because of runtime memory behavior. Limitación: Issue state and fixes can change; check the runtime version and referenced patch before relying on MLX.
  • GitHub · MLX cache/residency failure on an M3 Ultra 512GBMac Studio · M3 Ultra 512GB · 512GB · MLX-LMReports failure around 11,456–11,488 generated tokens in a production-style run · high confidenceEven very large unified-memory systems are not automatically stable for long-context MLX runs. Limitación: This is a runtime bug report, not a capacity limit; a patched release may change the result.
  • GitHub · DeepSeek V4 Flash Q8 garbled output on Apple NEONApple Silicon CPU / NEON · llama.cpp · Q8A specific repack path produced garbled output · high confidenceThis narrows the risk to a quantization/repack/runtime combination instead of treating every llama.cpp path as equivalent. Limitación: Do not generalize one broken Q8 repack to all DeepSeek V4 Flash formats.
  • GitHub · ds4 Apple Silicon throughput tableMac Studio · M3 Ultra 512GB · 512GB · ds4 Metal/CUDA engine · Q4 classRepository table reports short/long generation figures around 78.95 / 35.50 tok/s · medium confidenceA purpose-built engine provides a useful upper-bound signal for a large-memory Mac Studio, but it is not a mainstream runtime guarantee. Limitación: Single graph worker, no batching, and engine-specific optimizations; compare only directionally.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
DeepSeekLímiteDeepSeek-V4-Pro 1.6T-A49BDeepSeek V4 · 1600B total / 49B activeMLX-4bit · MLX960.0 GB · Public model source; exact file not registeredNo cabe-952GB restantes tras la reservaDesconocidollama.cpphighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public model source; exact file not registered960.0 GB · parameter planning estimate
  • Fuente del runtimellama.cpp · Desconocido
  • No se ha capturado un informe fiable de primera mano sobre Mac en esta actualización.La capacidad sigue siendo una estimación de planificación basada en metadatos públicos; la ausencia de pruebas no valida el runtime.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
GLMPrincipalGLM-5.3-Flash 320B-A18BGLM-5.3-Flash · 320B total / 18B activeMLX-4bit · MLX204.0 GB · Public file metadataNo cabe-195.99GB restantes tras la reservaInforme comunitariomlx2 señales sociales · ver fuenteshighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public file metadata204.0 GB · exact public file metadata
  • Fuente del runtimemlx · Informe comunitario
  • Evidencia de la comunidad (2)Los informes públicos son solo señales; nunca convierten esta página en una prueba del proyecto.
  • Hugging Face · Community GLM-5.3-Flash MLX quantization ladderMLX · 2bit-lite / 2 / 3 / 4 / 6-bit MLXPublic metadata lists about 102.43GB to 295.60GB per variant; the 4-bit root mirror is about 203.99GB · medium confidenceThe newer repository makes the precision trade-off explicit and includes a 2bit-lite path, but it does not prove loadability or speed on a particular Mac. Limitación: Community conversion; full-repository size includes duplicate root and variant folders, and no trustworthy Apple first-person benchmark is attached.
  • Reddit · GLM-5.3 Flash comparison thread (non-Apple hardware)Dual DGX Spark / other non-Apple systemsOne report describes roughly 22 tok/s on dual Spark and a slower, more verbose interaction style · low confidenceThis indicates active community comparison but cannot be translated into a Mac recommendation. Limitación: Cross-hardware, anecdotal, and workload-specific.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
GLMLímiteGLM-5.3 744B-A40BGLM-5.3 · 744B total / 40B activeMLX-4bit · MLX446.4 GB · Public model source; exact file not registeredNo cabe-438.4GB restantes tras la reservaDesconocidollama.cpphighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public model source; exact file not registered446.4 GB · parameter planning estimate
  • Fuente del runtimellama.cpp · Desconocido
  • No se ha capturado un informe fiable de primera mano sobre Mac en esta actualización.La capacidad sigue siendo una estimación de planificación basada en metadatos públicos; la ausencia de pruebas no valida el runtime.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
Google GemmaPequeñoGemma 4 E2BGemma 4 · 2.3B effective / 5.1B total · PLEQ4_0 · QAT3.35 GB · Public file metadataCon margen4.65GB restantes tras la reservaInforme comunitariollama.cpp1 señal social · ver fuenteshighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public file metadata3.35 GB · exact public file metadata
  • Fuente del runtimellama.cpp · Informe comunitario
  • Evidencia de la comunidad (1)Los informes públicos son solo señales; nunca convierten esta página en una prueba del proyecto.
  • Tech media / blog · Apple Silicon local-AI comparison including Gemma 4Mac Studio / M4 Max 128GB test systems · 128GB · llama.cppGemma 4 appears in the comparison; the article warns that bandwidth is not the sole predictor and does not publish this exact matrix tuple · low confidenceUseful context for Apple Silicon behavior, but not an exact Gemma 4 E2B fit or performance claim. Limitación: Different model/quantization details and media test methodology from this product matrix.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
Google GemmaPrincipalGemma 4 12B UnifiedGemma 4 · 11.95B denseQ4_0 · QAT6.98 GB · Public file metadataAjustado1.02GB restantes tras la reservaDesconocidollama.cpphighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public file metadata6.98 GB · exact public file metadata
  • Fuente del runtimellama.cpp · Desconocido
  • No se ha capturado un informe fiable de primera mano sobre Mac en esta actualización.La capacidad sigue siendo una estimación de planificación basada en metadatos públicos; la ausencia de pruebas no valida el runtime.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito
Google GemmaLímiteGemma 4 26B-A4BGemma 4 · 25.2B total / 3.8B activeQ4_0 · QAT14.4 GB · Public file metadataNo cabe-6.44GB restantes tras la reservaInforme comunitariollama.cpphighComprobado 2026-08-28
Fuentes y límites
  • Fuente oficial del modeloModelo family, parameters, and release context · Comprobado 2026-08-28
  • Public file metadata14.4 GB · exact public file metadata
  • Fuente del runtimellama.cpp · Informe comunitario
  • No se ha capturado un informe fiable de primera mano sobre Mac en esta actualización.La capacidad sigue siendo una estimación de planificación basada en metadatos públicos; la ausencia de pruebas no valida el runtime.
  • Public file metadata or a parameter estimate is not a first-party benchmark.
  • Community artifacts and reports do not imply vendor endorsement or Keep or Upgrade testing.
  • Social evidence is directional: build, context, thermals, storage streaming, and prompt mix can change the result.
  • Context, concurrency, tool calling, JSON, vision, and sustained behavior remain separate capability checks.
  • The primary matrix prioritizes current flagship releases; legacy versions remain compatibility records but are intentionally omitted here.
Usar en asesor gratuito

Evidencia de capacidad

La capacidad del flujo de trabajo es una cuestión aparte

Las ventanas de contexto son metadatos del modelo. La concurrencia, la llamada de herramientas, la confiabilidad de JSON, la visión y el comportamiento sostenido dependen del modelo exacto, el tiempo de ejecución, el cliente, la presión de la memoria y la tarea. Una celda positiva todavía no es garantía para su flujo de trabajo.

ModeloContextoConcurrenciaLlamadas a herramientasJSONVisiónUso sostenido
Qwen3.8 27BPequeño
Soporte oficialPublic model metadata lists 262,144 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe current Qwen material documents tool or agent paths. The exact model template and runtime still need a workflow check.Fuente
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
Qwen3.8-Flash-Next 180B (6B active)Principal
Soporte oficialPublic model metadata lists 262,144 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe current Qwen material documents tool or agent paths. The exact model template and runtime still need a workflow check.Fuente
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
Soporte oficialThe current model documentation describes multimodal inputs for this representative; the selected Mac runtime path still needs checking.Fuente
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
Qwen3.8 2.4T-A95BLímite
Soporte oficialPublic model metadata lists 262,144 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe current Qwen material documents tool or agent paths. The exact model template and runtime still need a workflow check.Fuente
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
DeepSeek-V4-Flash 284B-A13BPrincipal
Soporte oficialPublic model metadata lists 1,048,576 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe official DeepSeek V4 material documents tool use in the model family. The exact local template and runtime still need a workflow check.Fuente
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
DeepSeek-V4-Pro 1.6T-A49BLímite
Soporte oficialPublic model metadata lists 1,048,576 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe official DeepSeek V4 material documents tool use in the model family. The exact local template and runtime still need a workflow check.Fuente
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
GLM-5.3-Flash 320B-A18BPrincipal
Soporte oficialPublic model metadata lists 1,000,000 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe official GLM-5.3 material documents function calling and agent capabilities. It is not a Mac runtime integration benchmark.Fuente
Benchmark públicoPublic GLM evaluations cover agentic or structured tasks, but they do not establish JSON validity in every client.Fuente
Soporte oficialThe current model documentation describes multimodal inputs for this representative; the selected Mac runtime path still needs checking.Fuente
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
GLM-5.3 744B-A40BLímite
Soporte oficialPublic model metadata lists 1,000,000 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
Soporte oficialThe official GLM-5.3 material documents function calling and agent capabilities. It is not a Mac runtime integration benchmark.Fuente
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
Gemma 4 E2BPequeño
Soporte oficialPublic model metadata lists 131,072 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
DesconocidoNo model-specific public tool-calling result is strong enough for a positive claim in this snapshot.
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
Soporte oficialThe current model documentation describes multimodal inputs for this representative; the selected Mac runtime path still needs checking.Fuente
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
Gemma 4 12B UnifiedPrincipal
Soporte oficialPublic model metadata lists 262,144 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
DesconocidoNo model-specific public tool-calling result is strong enough for a positive claim in this snapshot.
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.
Gemma 4 26B-A4BLímite
Soporte oficialPublic model metadata lists 262,144 tokens. This is not proof that the selected Mac can hold that context with the weights.Fuente
Inferencia del sistemaConcurrency is inferred only from memory budget and runtime architecture. No apples-to-apples public concurrency benchmark is attached.Fuente
DesconocidoNo model-specific public tool-calling result is strong enough for a positive claim in this snapshot.
DesconocidoStructured JSON behavior depends on the exact model template, runtime parser, and client; no exact public result is registered here.
DesconocidoVision support is not assumed from family branding or an unrelated model variant.
DesconocidoNo public, comparable sustained Apple Silicon run is registered for this model and Mac configuration.

Cómo leer las etiquetas

  • Soporte oficial La documentación oficial del modelo/runtime describe la ruta de capacidad; el comportamiento exacto del cliente aún debe comprobarse.
  • Benchmark público Existe una evaluación publicada, pero no es un benchmark de rendimiento o flujo de trabajo en Mac.
  • Informe comunitario Los informes públicos o artefactos comunitarios son señales útiles, no pruebas del proyecto ni garantías del proveedor.
  • Inferencia del sistema Inferencia acotada a partir de la arquitectura de capacidad/runtime, no un resultado observado.
  • Desconocido No se ha capturado evidencia pública fiable para esta afirmación exacta.
  • Verificado por el proyecto Solo cuenta la combinación exacta de modelo, cuantización, runtime y máquina del registro de pruebas del proyecto.

Trae tu verdadera carga de trabajo

Convierta la evidencia pública en una decisión personal.

El Asesor gratuito agrega su Mac actual, contexto, simultaneidad, presupuesto, compatibilidad y requisitos fuera de línea. Puede decir “necesita verificación” cuando la evidencia pública no es suficiente.

Ejecute el asesor gratuito

Preguntas frecuentes

Verifique los límites de prueba antes de comprar.

¿Es este un punto de referencia de Mantener o Actualizar?

No. Una herramienta de planificación y síntesis de evidencia pública. Solo se mostrará el conjunto correcto de humo de proyectos como proyecto verificado.

¿Cómo se calcula la capacidad de tu Mac?

Utilice el tamaño del archivo público seleccionado si el archivo exacto o el total del archivo dividido está disponible; de ​​lo contrario, utilice la estimación de planificación explícita. El presupuesto modelo reserva espacio para macOS, el tiempo de ejecución y el contexto.

¿MoE requiere menos memoria si es un parámetro válido?

No. Los pesos completos almacenados son la condición de capacidad. Los parámetros válidos solo describen la cantidad de cálculo por token.

¿La longitud del contexto oficial funciona en cualquier Mac?

No. La longitud del contexto son los metadatos del modelo. Después del peso, el tiempo de ejecución, el sistema y otras aplicaciones, también necesitas tener suficiente memoria en tu Mac seleccionado.

¿Se puede comprar un Mac con sólo mirar la tabla?

Después de reducir el alcance de su planificación con esta matriz, ingrese su Mac actual, contexto, simultaneidad, clientes, presupuesto, compatibilidad y requisitos fuera de línea en nuestro Asesor gratuito. La evidencia pública no es garantía de compra.

traer el trabajo real

Utilice evidencia pública para informar sus decisiones de compra.

Free Advisor agrega su Mac actual, contexto, simultaneidad, presupuesto, compatibilidad, requisitos fuera de línea e indica que se requiere verificación si no hay evidencia pública suficiente.

Iniciar asesor gratuito

Keep or Upgrade es una herramienta independiente; Apple no patrocina ni respalda esta página.