← Back to tags

Tag

#agents

20 articles

NEW Model News

Tencent Hy4 Preview: 770B MoE, 1M Context, Pricing and Benchmarks

Tencent Hy4 preview checked: 770B/49B MoE, 1M context, API limits, pricing, Apache 2.0, self-hosting hardware and the evidence behind its benchmarks.

Modell-News

DeepSeek V4 Pro 0813: Preise, Benchmarks und Mac-Praxischeck

DeepSeek V4 Pro 0813 erklärt: API-Preise, Agent-Benchmarks, Pro-vs.-Flash-Vergleich und ob das 1,6-T-Modell auf einem Mac sinnvoll läuft.

Model News

DeepSeek V4 Pro 0813: Pricing, Benchmarks and the Mac Reality Check

DeepSeek V4 Pro 0813: API pricing, agent benchmarks, Pro vs Flash, API caveats and why the 1.6T model is not a realistic local Mac model.

Model News

Qwen3.8-27B is here: what the new 27B open-weight release means for local Macs

Qwen3.8-27B landed August 14, 2026 as an official open-weight checkpoint. Metadata, MLX/GGUF ecosystem, Mac memory math and remaining benchmark gaps.

Model News

Gemini 3.7 Flash on Mac: API Pricing, Benchmarks & Local Limits

Gemini 3.7 Flash is GA as of August 13, 2026. What Mac users need to know about API pricing, 1M context, coding benchmarks, privacy and local inference.

Model News

Kimi K3 on Mac: Open weights are here — why local use is still impractical

Kimi K3 for Mac: 2.8T parameters, 1M context, published open weights, API pricing and why full local inference still exceeds a normal Mac's hardware.

Cloud AI

Meta Muse Spark 1.1 on Mac: Current Status After Muse Spark 1.2

Muse Spark 1.1 is a hosted agent model, not a local Mac model. This guide separates 1.1 from Muse Spark 1.2 and explains API access, cost and data flow.

Model News

Grok on Mac: Grok Bot Desktop App, Grok Build and API

Looking for Grok on Mac? Compare the Grok Bot desktop app, Grok Build and the Grok 4.5 API: availability, setup, cloud limits and local alternatives.

Modell-News

Tencent Hy3 auf dem Mac: OpenRouter, 295B MoE, Apache 2.0 und lokale Grenzen

Tencent Hy3 erklärt: 295B MoE, 21B aktive Parameter, 256K Kontext, OpenRouter-Slug tencent/hy3 und warum lokale Mac-Inferenz unrealistisch bleibt.

Model News

Tencent Hy3 on Mac: OpenRouter, 295B MoE, Apache 2.0 and Local Limits

Tencent Hy3 explained: 295B MoE, 21B active parameters, 256K context, OpenRouter slug tencent/hy3 and why local Mac inference stays unrealistic.

Model News

Claude Sonnet 5 on Mac: Agents, Coding, 1M Context and API Costs Explained

Claude Sonnet 5 explained: API pricing, 1M context, Claude Code, agent workflows, model IDs and why it runs in the cloud instead of locally on Mac.

Cloud AI

Sakana Fugu Ultra: An AI Orchestrator, Not a Model You Can Download

Sakana Fugu Ultra is not a local LLM but a cloud orchestrator that coordinates multiple models. What that means for Mac users, EU availability, and pricing.

Cloud AI

Kimi K2.7 Code on Mac: Cloud, API, or GGUF?

Kimi K2.7 Code on Mac, separated clearly: Ollama Cloud, Kimi Code, API pricing, mandatory thinking, and the demanding GGUF route.

Model News

Claude Fable 5 Is Back: Status, Pricing and Mac Alternatives

Anthropic is redeploying Claude Fable 5 after US export controls were lifted. Current status for Claude Code, API, pricing, data retention and Mac alternatives.

Cloud AI

NVIDIA Nemotron 3 Ultra on Mac: Cloud Tag, 256K and NIM

Use Nemotron 3 Ultra from a Mac without confusing cloud access with local inference: Ollama Cloud, native 256K, optional NIM 1M, real hardware limits.

Cloud AI

StepFun Step 3.7 Flash on Mac: 198B MoE, 256K Context and the Local Reality

StepFun Step 3.7 Flash explained: 198B MoE, 11B active parameters, 256K context, API pricing and why normal Macs are not enough locally.

Cloud-KI

MiniMax M2.7 auf dem Mac: Cloud-API, Token Plan und lokale Grenzen

MiniMax M2.7 sachlich erklärt: Modellleistung, API- und Token-Plan-Preise und was die offiziellen Quellen über lokale Mac-Nutzung sagen.

Cloud AI

MiniMax M2.7 on Mac: Cloud API, Token Plan, and Local Limits

MiniMax M2.7 on Mac: capabilities, API and Token Plan pricing, Ollama Cloud, and what official sources say about local use.

Modell-News

Qwen3.7 Max: Lohnt sich OpenRouter?

Qwen3.7 Max auf OpenRouter: Tokenpreise, 1M Kontext, API-Setup und warum das Modell in der Cloud statt lokal auf dem Mac läuft.

Model News

Qwen3.7-Max OpenRouter Pricing: 1M Context, API Setup & Mac Limits

Qwen3.7 Max on OpenRouter: current token pricing, 1M context, API setup and why the model runs in the cloud rather than locally on a Mac.