Skip to content
01 · Scope
02 · Build
03 · Ship
3hship · a community for reading the blog and AI news

How much can you ship in 3 hours?

Read the 3hship blog, follow the AI news worth your time and talk it through with other readers. New posts reach you in the newsletter and on Discord.

From the blog

  1. Evaluating local models on one 12 GB GPU

    Five local models tested on one RTX 4070 (12 GB) with LM Studio and llama.cpp: Qwen3-VL-8B, Gemma 4 12B QAT, Gemma 4 26B A4B, Qwen3.8-27B and Qwen3.6-35B-A3B. Accuracy against 40 hand-labelled documents, time to first token, tokens per second, MoE expert offload tuning (1.1 to 52.8 tok/s), prompt-size limits and vision, thinking and tool-calling checks, as a method to repeat on your own card.

    8 min read

All posts

Join 3hship on Discord

Read the blog, follow the AI news and talk it through with the other readers.

What is in the server

  • #shipped_chat: post what you shipped
  • #build_live: work in public while a block runs
  • Help forum: get unstuck
  • #general: everything else
Join the Discord

The public invite link is not published yet. The newsletter will carry it when it is.