Qwen 3.8 27B (Fully Tested on Local Mac 48GB) + Hermes: I might CANCEL Claude & Codex NOW!

AICodeKing18m 13s
0 comments · 0 votesOpen discussionClose discussion
Sign in to join the discussion

    Video summary

    Qwen 3.8 27B is framed as a local agent model rather than a frontier reasoning model. The video focuses on its long context, native vision support and tool calling, while warning that harder reasoning tasks still favor larger hosted systems.

    The presenter compares reported coding and computer use benchmarks, then explains how reasoning effort settings affect speed and depth. These scores are claims shown in the source and were not independently verified for this draft.

    For local serving, the walkthrough recommends Ollama for the simplest setup and LM Studio for more control over quantization, GPU offload and key value cache memory. It emphasizes increasing the context window and checking tool use settings.

    Hermes Agent sits above the local endpoint to provide orchestration, persistent memory, subagents and integrations. The proposed stack is best suited to long, routine sequences such as reading files, running commands and navigating interfaces, not difficult architecture or debugging decisions.

    Original YouTube thumbnailWatch on YouTube