← Back to tags

Tag

#local-ai

24 articles

NEU Modell-News

Qwen3.8-Flash im Faktencheck: Preise, 1M Kontext, Benchmarks und Flash-Next erklärt

Qwen3.8-Flash geprüft: aktuelle Preise, 1M Kontext, API-Funktionen, Benchmarks und der wichtige Unterschied zu Qwen3.8-Flash-Next.

NEW Model News

Qwen3.8-Flash Explained: Pricing, 1M Context, Benchmarks, and Flash-Next

A source-checked guide to Qwen3.8-Flash pricing, its 1M context window, API features, benchmarks, and the separate Flash-Next open-weight release.

Model News

Qwen3.8-27B is here: what the new 27B open-weight release means for local Macs

Qwen3.8-27B landed August 14, 2026 as an official open-weight checkpoint. Metadata, MLX/GGUF ecosystem, Mac memory math and remaining benchmark gaps.

Modell-News

Gemini 3.7 Flash auf dem Mac: API-Preis, Benchmarks & lokale Grenzen

Gemini 3.7 Flash ist seit 13. August 2026 GA. Was Mac-Nutzer über API-Preis, 1M-Kontext, Coding-Benchmarks, Datenschutz und lokale Nutzung wissen müssen.

Model News

Gemini 3.7 Flash on Mac: API Pricing, Benchmarks & Local Limits

Gemini 3.7 Flash is GA as of August 13, 2026. What Mac users need to know about API pricing, 1M context, coding benchmarks, privacy and local inference.

Local Models

Qwen3.6-27B Fable Fusion 711 on Mac: GGUF Sizes and RAM Choices

Which Fable Fusion 711 GGUF fits 24, 32, 48 or 64 GB of unified memory? File sizes, MTP, vision and benchmark evidence—without claiming original Mac benchmarks.

Local Models

Apple Intelligence Local AI: On-Device Models, PCC and Apple Silicon

Learn what Apple Intelligence runs on-device, when Private Cloud Compute is used, and how Apple Foundation Models work on Apple silicon.

Modell-News

Kimi K3 auf dem Mac: Open Weights sind da – warum lokal trotzdem unpraktisch bleibt

Kimi K3 auf dem Mac: 2,8T Parameter, 1M Kontext, veröffentlichte Gewichte, API-Preise und warum lokale Inferenz einen normalen Mac überfordert.

Model News

Kimi K3 on Mac: Open weights are here — why local use is still impractical

Kimi K3 for Mac: 2.8T parameters, 1M context, published open weights, API pricing and why full local inference still exceeds a normal Mac's hardware.

Modell-News

Grok auf dem Mac: Grok Bot, Grok Build und API

Du suchst Grok für den Mac? Vergleiche die Grok-Bot-Desktop-App, Grok Build und die Grok-4.5-API: Verfügbarkeit, Setup, Cloud-Grenzen und lokale Alternativen.

Model News

Grok on Mac: Grok Bot Desktop App, Grok Build and API

Looking for Grok on Mac? Compare the Grok Bot desktop app, Grok Build and the Grok 4.5 API: availability, setup, cloud limits and local alternatives.

Modell-News

Poolside Laguna XS.2 auf dem Mac: Open-Weight Coding-Modell, Benchmarks und RAM

Poolside Laguna XS.2 ist ein offenes 33B-MoE-Coding-Modell mit 3B aktiven Parametern. Benchmarks, Fähigkeiten und realistische Mac-Konfigurationen.

Model News

Poolside Laguna XS.2 on Mac: Open-Weight Coding Model, Benchmarks and RAM

Can Poolside Laguna XS.2 run on a Mac? See RAM needs, coding benchmarks, Ollama options and which Apple Silicon Macs fit the 33B MoE model.

Modell-News

Claude Sonnet 5 auf dem Mac: Agenten, Coding, 1M Kontext und API-Kosten erklärt

Claude Sonnet 5 erklärt: 1M Kontext, adaptive Thinking, Preise und Claude Code – und warum das Modell nicht lokal auf dem Mac läuft.

Model News

Claude Sonnet 5 on Mac: Agents, Coding, 1M Context and API Costs Explained

Claude Sonnet 5 explained: API pricing, 1M context, Claude Code, agent workflows, model IDs and why it runs in the cloud instead of locally on Mac.

Cloud-KI

GLM-5.2 auf dem Mac: OpenRouter, 1M-Kontext und Grenzen

GLM-5.2 von Z.ai erklärt: 1M-Kontext, OpenRouter-Setup, aktuelle Preisquellen und warum die dokumentierte Mac-Nutzung primär über die Cloud läuft.

Modell-News

Gemma 4 12B auf dem Mac: Das neue lokale Multimodal-Modell für 16 GB?

Gemma 4 12B läuft lokal ab 16 GB, bietet 256K Kontext sowie Bild- und Audioverständnis. Was auf dem Mac mit Ollama und MLX wirklich geht.

Model News

Gemma 4 12B on Mac: Is 16 GB Really Enough?

Gemma 4 12B runs locally from 16 GB with 256K model context and multimodal input. What Ollama and MLX actually support on Mac.

Cloud AI

MiniMax M2.7 on Mac: Cloud API, Token Plan, and Local Limits

MiniMax M2.7 on Mac: capabilities, API and Token Plan pricing, Ollama Cloud, and what official sources say about local use.

Cloud AI

Gemini 3.5 Flash on Mac: How to Use It, Pricing, and Local Alternatives

What Gemini 3.5 Flash can do, how to use it from a Mac, what it costs, how Google handles your data, and which local models fit when you need offline AI.

Modell-News

Qwen3.7 Max: Lohnt sich OpenRouter?

Qwen3.7 Max auf OpenRouter: Tokenpreise, 1M Kontext, API-Setup und warum das Modell in der Cloud statt lokal auf dem Mac läuft.

Model News

Qwen3.7-Max OpenRouter Pricing: 1M Context, API Setup & Mac Limits

Qwen3.7 Max on OpenRouter: current token pricing, 1M context, API setup and why the model runs in the cloud rather than locally on a Mac.

Comparisons

Apple Intelligence vs Local AI: Mac Privacy Guide

Apple Intelligence, PCC, ChatGPT and local AI on Mac: what stays local, when cloud processing happens and when Ollama is more private.

Vergleiche

Apple Intelligence vs. lokale KI: Datenschutz auf dem Mac

Apple Intelligence, PCC, ChatGPT und lokale KI auf dem Mac: Welche Daten lokal bleiben, wann Cloud greift und wann Ollama privater ist.