HyphenTech · All posts

Back to home

Page 2 of 3 · 21 posts per page. There's more to read. Glad you're here.

People say it's faster than Ollama, but no one has actually competed with her

Ollama switched Mac inference to MLX kernel half a year ago, and now a bunch of challengers claiming to be faster have emerged. I didn't run the tests, just put the scores published by the four companies on one table—the result was: none of them were running the same set of problems.

GPT-6 is here—is it really starting to work for you?

The real change in GPT-6 Astra isn't just a slight increase in chat score, but the simultaneous integration of computer operation, long tasks, and high-risk capabilities into the product. This article breaks down the available range, API costs, security boundaries, and how ordinary people choose based on OpenAI's official materials.

8 GB VRAM can run local AI: How to choose between Windows and Mac?

Instead of discussing superficial questions like "Is there a Windows version of the software?", let's answer two practical questions: which local models and tools can run on an 8 GB NVIDIA graphics card, and what makes the Mac's unified memory and complete multimedia chain strong.

DeepSeek Visual Open Source—Can a 64GB Mac Run It?

DeepSeek turns vision into the Agent's perception entry point and opens up a 305B weight. The latest minimum visual GGUF is still about 67.8GiB: 64GB Mac is not suitable for stable deployment today. This article is based on official data and files, and does not use local running impersonation.

Fable 5.1 didn't drop in price, so why save 45%?

Fable 5.1 and Mythos 5.1 are the same underlying model, but were split into public and trusted access versions. The base unit price hasn't dropped, but cache reads are 75% cheaper; What truly changes are the task economics, security boundaries, and access methods for long-term agents.

LocalBrain: turn your Mac into a private AI box

I gathered the local models, transcription, image and video generation, documents and MCP scattered across my Mac into one local workbench that downloads, starts, chats, delivers files and cleans up its own caches. 1.4.8 adds a second local video engine, LTX-2.5, with a direct download from mainland China on Discover; a full tour with screenshots.

Completely free: Get 103 large model APIs for free | HyphenBox

HyphenBox has officially launched, starting from version 1.0.0, and offers installation packages for Mac, Windows, and Linux. Early screenshots and detection numbers in this article are retained as development records and do not represent the current free quota.

The mysterious model exploded! Niu Lai?

Late at night on August 20, OpenRouter added a line of stealth/ox-alpha. No company, no warehouse, no name, but in two days it shot to number one in call volume. Around 1.04 million yuan, you can watch videos, completely free—the price is written at the top of the page.

Mac Local Raw Video: Official Introduction to FastMetal-QAD

The FastVideo team has launched three open-source video models for Apple Silicon: 1.3B, 5B, and 14B. This article compiles memory thresholds, speed gauges, generation modes, and installation entry points based on official materials, excluding local tests.